arXiv AI By Zhimin Hu, Jeroen van Paridon, Gary Lupyan

Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Language

Read the original on arXiv AI →

arXiv:2607. 04523v1 Announce Type: cross Abstract: Generic statements like "tigers are striped" and "cars have radios" communicate information that is, in general, true.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

Hugging Face Trending Papers
Jul 8

Dissociating the Internal Representations of Sycophancy in LLMs

Large Language Models (LLMs) frequently exhibit sycophancy, where they agree with a user's statement even when incorrect. While sycophancy is often treated as a single defined behavior, it can manifest in substantially distinct ways and circumstances, raising the question of whether this multi-faceted nature is reflected in its internal mechanisms.

Hugging Face Trending Papers
Jun 11

Reasoning as Pattern Matching: Shared Mechanisms in Human and LLM Everyday Reasoning

When large language models (LLMs) fail to generalize or make haphazard errors in reasoning, it is often taken as evidence that LLMs are not truly reasoning, but rather performing a kind of pattern matching. The implication is that people's behavior does not exhibit the same types of failures because human reasoning uses principled and abstract world models.

arXiv AI
2d ago

Verbalized and Internal Probabilities Are Coupled in Large Language Models

The paper investigates the relationship between a large language model’s internal probability distribution and its verbalized confidence statements. By systematically manipulating training and in‑context data, the authors show that both internal and verbalized probabilities are influenced by distributional and asserted uncertainty in the data. They find that verbalized probabilities align with internal ones beyond what would be expected if they tracked the same sources independently, indicating that verbalized confidence can serve as a probe of the model’s internal distribution.

By Sinead Williamson, Jiaxuan Li, Nick Foti, Russ Webb, Masha Fedzechkina