arXiv AI By Ziyang Ling, Ronald X. Xu, Mingzhai Sun

THOR: A Theta-Gamma Hierarchical Oscillatory Reasoning Framework for Multi-hop QA

Read the original on arXiv AI →

arXiv:2607. 20459v1 Announce Type: cross Abstract: Multi-hop question answering requires retrieving and integrating evidence from multiple contexts.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
Sep 22

SG-FSM: A Self-Guiding Zero-Shot Prompting Paradigm for Multi-Hop Question Answering Based on Finite State Machine

arXiv:2410.17021v2 Announce Type: replace Abstract: Large Language Models with chain-of-thought prompting, such as OpenAI-o1, have shown impressive capabilities in natural language inference tasks. H...

By Xiaochen Wang, Liang Chen, Reza Haf Zhe Yang, Yiru Wang, Xiangdi Meng, Kunhao Pan, Zhifang Sui, Junqing He
arXiv AI
Oct 1

Hierarchical Reasoning Model

arXiv:2506.21734v4 Announce Type: replace Abstract: Reasoning, the process of devising and executing complex goal-oriented action sequences, remains a critical challenge in AI. Current large language...

By Guan Wang, Jin Li, Yuhao Sun, Xing Chen, Changling Liu, Yue Wu, Meng Lu, Sen Song, Yasin Abbasi Yadkori
arXiv AI
Sep 30

Rethinking Reasoning Paths as Phase-Structured Trajectories

The paper proposes PAIR, a method that treats reasoning paths of large language models as phase‑structured trajectories within each question. By sampling multiple trajectories per question, aligning them to shared relative phases, and comparing successful versus unsuccessful paths only within the same phase, PAIR isolates path‑quality signals from question‑level variation. Experiments show that standard correctness probes lose predictive power under this within‑question evaluation, while PAIR improves trajectory ranking, Best‑of‑N selection, and enables phase‑wise steering of generation outcomes.

By Zhenghao He, Guangzhi Xiong, Sanchit Sinha, Bohan Liu, Wenqian Ye, Aidong Zhang
arXiv AI
Sep 24

The Path Matters: Evaluating Small Language Models Beyond Answer Accuracy in KGQA

The paper investigates how small language models (SLMs) perform in knowledge graph question answering (KGQA) when evaluated on the reasoning paths they take, rather than just the final answer. Using the THESEUS navigation and traceability framework, the authors test frozen, off‑the‑shelf SLMs as local action policies that choose graph actions and decide when to stop, without any task‑specific training or free‑form answer generation. By measuring both Hits@1 and Path Edit Distance (PED) across the Kinship and MQuAKE‑ST datasets, the study finds that models vary significantly in both answer accuracy and path fidelity, and that prompting can either help or hurt navigation depending on the model. "whyItMatters":"The results show that evaluating SLMs solely on endpoint accuracy can be misleading, highlighting the need to assess reasoning path fidelity in KGQA tasks."

By Eduin E. Hernandez, Sergio A. Diaz, Luis F. Garcia, Nurassyl Askar, Stefano Rini