← Back to all news
Hugging Face Trending Papers June 8, 2026

Multi-Hop Knowledge Composition is Bound by Pretraining Exposure

Read the original on Hugging Face Trending Papers →

Large Language Models fail at implicit multi-hop reasoning: a model answers "When was $X$ born? " and "Who is $Y$'s closest friend?

Summary generated by The Flow from the publisher's feed. The full article lives at Hugging Face Trending Papers.

  • llms

Related stories

arXiv AI
Jul 2

DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoning

arXiv:2607. 00341v1 Announce Type: cross Abstract: Large language models achieve strong performance on many reasoning tasks when allowed to externalize intermediate steps as Chain-of-Thought (CoT).

By Hengyu Fu, Tianyu Guo, Zixuan Wang, Hanlin Zhu, Jason D. Lee, Jiantao Jiao, Stuart Russell, Song Mei
llmsragbenchmarks
More like this →
arXiv AI
Aug 12

Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers

arXiv:2604. 07822v2 Announce Type: replace-cross Abstract: We study implicit reasoning, i.

By Harsh Kohli, Srinivasan Parthasarathy, Huan Sun, Yuekun Yao
llms
More like this →
arXiv AI
Jun 2

Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models

arXiv:2504. 03635v4 Announce Type: replace Abstract: Reasoning is a core capability of language models (LMs), yet it remains unclear how much model capacity is necessary to support reasoning during pretraining.

By Xinyi Wang, Shawn Tan, Shenbo Xu, Mingyu Jin, William Yang Wang, Rameswar Panda, Yikang Shen
llms
More like this →
arXiv AI
Jul 10

ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation

arXiv:2502. 15543v4 Announce Type: replace-cross Abstract: Large language models (LLMs) integrated with retrieval-augmented generation (RAG) have improved factuality by grounding outputs in external evidence.

By Pengcheng Huang, Zhenghao Liu, Yukun Yan, Haiyan Zhao, Xiaoyuan Yi, Hao Chen, Zhiyuan Liu, Maosong Sun, Tong Xiao, Ge Yu, Chenyan Xiong
llmsragbenchmarks
More like this →
arXiv AI
Jul 14

Unveiling the Mechanisms of Multi-Hop Reasoning in Transformers via Identity Bridge

arXiv:2509. 24653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at multi-hop reasoning in distribution, yet fail on unseen compositions, a phenomenon known as the curse of two-hop reasoning.

By Pengxiao Lin, Zheng-An Chen, Zhi-Qin John Xu
llmsfine-tuning
More like this →
arXiv AI
Jun 2

Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge

arXiv:2602. 02470v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) have achieved remarkable success in many complex tasks, yet they can still fail in very simple logical reasoning such as the "reversal curse" -- when trained on forward knowledge data of the form "$A \rightarrow B$" (e.

By Xutao Ma, Yixiao Huang, Hanlin Zhu, Somayeh Sojoudi
llmssafety
More like this →