← Back to all news
Hugging Face Trending Papers July 27, 2026

Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models

Read the original on Hugging Face Trending Papers →

Appending a two-word confirmation tag to a decision question -- "Is X the better choice? " versus "X is the better choice, right?

Summary generated by The Flow from the publisher's feed. The full article lives at Hugging Face Trending Papers.

  • llms
  • rag
  • benchmarks

Related stories

arXiv AI
Jul 28

Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models

arXiv:2607. 23976v1 Announce Type: cross Abstract: Appending a two-word confirmation tag to a decision question -- "Is X the better choice?

By Tapan Parikh
llmsragbenchmarks
More like this →
arXiv Machine Learning
Jul 14

Two Confounds in Cross-Model Value Comparison: Response Determinism and the Access Harness

arXiv:2607. 10202v1 Announce Type: new Abstract: Cross-model comparisons read divergence in value dispositions as evidence that language models hold individuated values.

By Hong-In Won, Jinseok Jang, Hyoseop Kim
llmsagents
More like this →
arXiv AI
Jul 15

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

arXiv:2607. 12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer every other model picks?

By Tapan Parikh
llmsrag
More like this →
Hugging Face Trending Papers
Jul 14

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer every other model picks? Asked to "pick a word -- any word," 44 models chose "serendipity" 41% of the time.

llmsrag
More like this →
arXiv AI
Jul 21

Committed Before Reasoning: Behavioral Reproduction and Preliminary Activation-Level Evidence of Answer Pre-Commitment in an Open-Weight LLM

arXiv:2607. 16451v1 Announce Type: cross Abstract: Chat models sometimes commit to an answer and then produce reasoning that justifies it rather than deriving it -- even when the answer contradicts a task premise.

By Heejin Jo
llmssafety
More like this →
arXiv Machine Learning
Aug 11

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

arXiv:2607. 10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model "value dispositions," and a concentration/extremity index over repeated draws is read as how sharply a model commits.

By Hong-In Won, Jinseok Jang, Hyoseop Kim
llmsbenchmarkssafety
More like this →