arXiv Machine Learning

Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs

arXiv:2608. 05660v1 Announce Type: new Abstract: As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning has become an important practical problem.

arXiv Machine Learning
Sep 14

Scaling Online Complex Event Detection with Synthetic Supervision and Mamba-Based Neural Algorithmic Reasoning

The paper presents NAROCE, a Neural Algorithmic Reasoning framework for online complex event detection (CED). It decouples rule learning from sensor semantics by pretraining a Mamba-based rule reasoner on synthetic atomic event traces and then adapting it to raw sensor inputs with limited labeled data. Experiments on a simulator‑generated benchmark show that NAROCE matches or surpasses strong baselines while using far fewer labeled sequences and computational resources.

By Liying Han, Gaofeng Dong, Xiaomin Ouyang, Kang Yang, Lance Kaplan, Federico Cerutti, Mani Srivastava
arXiv AI
6d ago

Can Linguistic Reasoning Vectors Enhance Multimodal Reasoning Ability?

The paper introduces LIFT, a lightweight vector‑intervention technique that transfers reasoning capability from a base large language model (LLM) to a vision‑language model (VLM) without retraining the VLM backbone. LIFT defines Reasoning Vectors as differences in hidden states between a reasoning path with an explicit trace and a solver path without it, and injects these vectors into the VLM’s language‑side activations. Experiments on two VLMs across six reasoning benchmarks show that vectors derived from the base LLM consistently outperform those derived from the aligned VLM, indicating that the base LLM is a more effective source for recovering degraded reasoning. "whyItMatters":"The study demonstrates that a simple, frozen‑backbone intervention can partially restore reasoning abilities in multimodal models, highlighting the value of leveraging the original language model’s reasoning power."

By Ziyi Wang, Li Li, Aolin Zhou, Yankun Shen, Chonghan Liu, Shuxia Lin, Xu Yang
arXiv AI
3d ago

Soft Spatial Reasoning

Soft Spatial Reasoning introduces a post‑training framework for Large Vision‑Language Models that replaces hard, token‑by‑token chain‑of‑thought reasoning with a soft, continuous state formed by mixing token embeddings at each intermediate step. The method employs AdaptSoft, a controller that adjusts the degree of softness based on hidden states and predictive uncertainty, guided by a gradient‑alignment learning objective that requires no intermediate supervision. Experiments on diverse spatial benchmarks show that this approach outperforms both hard and fixed‑soft chain‑of‑thought baselines and several existing LVLMs.

By Rafi Ibn Sultan, Md. Sajid Alam Chowdhury, Saleh Zare Zade, Chengyin Li, Prashant Khanduri, Marco Brocanelli, Dongxiao Zhu
arXiv Computation and Language
Sep 11

RetroThinker: Enabling Retrospective Thinking in Speech LLMs

RetroThinker is a multi-stage post‑training framework that enhances SpeechLLMs by enabling them to self‑verify and forward‑correct Chain‑of‑Thought reasoning steps during inference. It combines supervised fine‑tuning on curated retrospective thinking data with length‑based direct preference optimization to improve reasoning while the user speaks. On the GSM8K benchmark, RetroThinker achieves an 11% absolute accuracy gain over non‑retrospective baselines while maintaining comparable latency.

By Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed, David Harwath