arXiv Machine Learning By Katharina Trinley, Jesujoba O. Alabi, Dietrich Klakow, Vagrant Gautam

A Mechanistic Understanding of Pronoun Fidelity in LLMs

Read the original on arXiv Machine Learning →

arXiv:2606. 16407v1 Announce Type: cross Abstract: Faithful and robust pronoun use is important for fair and coherent generations, yet large language models largely fail when multiple referents use different pronouns.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Computation and Language
Sep 4

No country for old linguists: LLM-brain alignment underdetermines neural computation

Nastase et al. (2026) argue that large language models (LLMs) can shed light on language processing because both use distributed, context‑sensitive representations shaped by statistical learning, and they advocate for LLM‑brain alignment research. They reject simple cortical “boxology” but claim that representational alignment can constrain mechanistic hypotheses, though it does not itself identify a mechanism. The author critiques this position, pointing out logical, causal, and computational underdetermination and the tension between the authors’ methodological caveats and their conclusion that LLMs could serve as fully mechanistic models of language.

By Elliot Murphy
Hugging Face Trending Papers
Jul 8

Dissociating the Internal Representations of Sycophancy in LLMs

Large Language Models (LLMs) frequently exhibit sycophancy, where they agree with a user's statement even when incorrect. While sycophancy is often treated as a single defined behavior, it can manifest in substantially distinct ways and circumstances, raising the question of whether this multi-faceted nature is reflected in its internal mechanisms.