arXiv AI

Comparing Semantic Navigation in Humans and Large Language Models using Natural Language Processing

arXiv:2607. 12195v1 Announce Type: cross Abstract: Semantic memory retrieval can be conceptualized as navigation through conceptual space.

arXiv Computation and Language
Aug 31

AI Models Can Predict and Collaboratively Modulate Human Memory Search

The study investigates how large language models (LLMs) can assist humans in semantic memory search tasks. By using the semantic fluency task (SFT), the researchers evaluate whether LLMs can follow and enhance human mental trajectories during generative semantic retrieval. Results show that an LLM’s ability to track and predict human memory trajectories in this task surpasses that of other humans.

By Eric Lacosse, Mariana Duarte, Graham Todd, Peter M. Todd, Daniel C. McNamee
arXiv AI
Aug 28

Artificial Intelligence Models Can Predict and Collaboratively Modulate Human Memory Search

The article reports that large language models can predict and collaboratively modulate human memory search during a semantic fluency task. By tracking and forecasting participants’ semantic retrieval patterns, the models outperform other humans in following these mental trajectories. This suggests that AI can serve as a cognitive tool to extend human abilities in open‑ended conceptual exploration and creative ideation.

By Eric Lacosse, Mariana Duarte, Graham Todd, Peter M. Todd, Daniel C. McNamee
arXiv Machine Learning
1d ago

Beyond Linear Concepts: Discovering and Aligning Non-Linear Concept Manifolds in Large Language Models

The paper extends mechanistic interpretability of large language models by modeling concepts as low‑dimensional non‑linear manifolds rather than linear subspaces. It introduces a concept‑based alignment (CBA) score to compare these manifolds across layers and models, revealing block structures in intermediate layers, a shift from syntax‑dominated to mixed syntactic‑semantic concepts, and training‑dependent multilingual sharing. The study also shows that alignment patterns differ across model families and training stages, with adjacent stages aligning more closely than distant ones.

By Tido Specht, Elias Benedict Krey, Nils Neukirch, Nils Strodthoff
arXiv AI
Sep 3

TUX: Measuring Human--AI Tacit Understanding

The paper introduces TUX, a Tacit Understanding Index that measures how similarly humans and large language models (LLMs) place concepts along subjective spectra in a task inspired by the game Wavelength. Using 241 human participants and 200 profile-conditioned LLM agents across four models, the study finds that human–agent pairs with similar traits achieve higher TUX scores, indicating that tacit alignment is linked to person-level characteristics. Regression analyses show that richer predictor sets—including individual traits, decision-making styles, and confidence—improve the explainability of TUX beyond simple trait-distance baselines.

By Yueshen Li, Hanyi Min, Vedant Das Swain, Koustuv Saha
arXiv Computation and Language
Sep 10

From Retrieval to Weights: Parametric Individualization of Small Language Models with Individual Text Corpora

The paper explores how individual text corpora (ITCs) can be integrated into small language models (SLMs) using DoRA fine‑tuning. By training a DoRA adapter for each of 150 participants, the authors show that the adapter can encode a participant’s own ITC into the model’s weights, improving fit to that participant’s held‑out text. However, while the adapter improves log‑loss on a generalized knowledge test, it does not enhance accuracy under a bias‑corrected PMI readout, and adding retrieval does not provide further benefit.

By Christoph Wigbels, Ali Abusaleh, Markus T. Jansen, Alexander Mehler, Markus J. Hofmann
arXiv Computation and Language
Aug 31

Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations

The paper investigates the intrinsic dimension (ID) of large language model (LLM) representations as an indicator of linguistic complexity. By comparing ID across model layers for coordination vs. subordination, right‑branching vs. center‑embedding, and unambiguous vs. ambiguous attachment, the authors find consistent ID differences that align with established complexity contrasts. Experiments across six LLMs, including representational similarity and layer pruning analyses, confirm that more complex phenomena produce higher ID profiles, with peaks occurring at different layers for each contrast.

By Marco Baroni, Emily Cheng, Iria de-Dios-Flores, Francesca Franzon
arXiv Computation and Language
4d ago

Cognitive Expert Language Models Better Align with the Corresponding Brain Systems

The study investigates whether language models tailored to specific cognitive domains better align with corresponding brain systems. By prompting and fine‑tuning large language models into six domain experts—sensory, spatial, numerical, reasoning, social, and abstract—the authors find that each expert’s representations more closely match the brain region associated with its domain than other experts. This domain‑specific alignment holds across multiple base models and fMRI datasets, while overall prediction accuracy remains largely unchanged, indicating that regional alignment can be obscured when summarizing across the brain.

By Zhivar Sourati, Mengxuan Helen Wu, Nona Ghazizadeh, Jonas Kaplan, Morteza Dehghani, Samuel A. Nastase