arXiv AI

LCoT-GV: Graph Attention Networks for Verifying Long Reasoning Chains in Large Language Models

arXiv AI
Aug 14

Unified Multi-Dimensional Benchmark for Complex Graph Reasoning in Large Language Models

arXiv:2608. 12391v1 Announce Type: cross Abstract: Graph reasoning provides a promising testbed for evaluating the reasoning ability of large language models (LLMs), as graph instances can be programmatically generated, structurally controlled, and naturally scaled to long-input settings.

By Fali Wang, Ali Al-Lawati, Iliyas Bektas, Jinxuan Fang, Alek Melenski, Tianxiang Zhao, Yao Ma, Suhang Wang
arXiv AI
Sep 2

KGFR: A Foundation Retriever for Generalized Knowledge Graph Question Answering

KGFR introduces a Knowledge Graph Foundation Retriever that collaborates with large language models to enhance knowledge‑intensive question answering. By encoding relations with LLM‑generated descriptions and initializing entities from question roles, KGFR enables zero‑shot generalization to unseen knowledge graphs. Its Asymmetric Progressive Propagation technique efficiently handles large graphs, while a controllable reasoning loop allows the LLM to request candidate answers, supporting facts, and reasoning paths.

By Yuanning Cui, Zequn Sun, Wei Hu, Zhangjie Fu
arXiv AI
Aug 5

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

arXiv:2608. 03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneous document elements distributed across multiple pages.

By Le Xiang, Zhicheng Guan, Hong Chen, Xiaocong Lin, Zhenghua Lei, Teng Hu, Bolei He, Long Zeng
arXiv Computation and Language
Sep 21

When Does Reasoning Help in Machine Translation? A Hierarchical Analysis of LRM Reasoning Traces

The paper investigates when intermediate reasoning traces benefit machine translation by examining models, languages, domains, and datasets. It finds that the optimal reasoning language depends on the model, reasoning length has a non‑monotonic effect on quality, and traces display recurring functional patterns. Using Hierarchical Meta‑Summarization, the authors uncover a shared structure of understanding/planning, translating/drafting, and refining/verifying, while also noting domain‑specific variations, suggesting that reasoning should be tailored rather than uniformly applied.

By Yuxiang Liu, Jiaming Luo, Eleftheria Briakou, Colin Cherry