arXiv Machine Learning By Ishan S. Kshirsagar

The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning

Read the original on arXiv Machine Learning →

arXiv:2607. 20952v1 Announce Type: new Abstract: Latent, or silent, reasoning lets language models carry out intermediate computation in continuous vector space instead of words, and is widely assumed to function as an internal scratchpad the model actively consults during inference.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 2

Latent Recurrent Thoughts: Recurrent Refinement of Proposed Latents for Reasoning with Frozen LLMs

Latent Recurrent Thoughts (LRT) proposes a method for reasoning with frozen large language models by operating in the model’s continuous representation space. A small auxiliary network generates initial latent vectors, which a tiny recurrent reasoner refines over multiple steps, decoupling computational depth from model size. Experiments on symbolic and natural‑language reasoning tasks show that LRT outperforms prior frozen‑decoder continuous‑space methods and chain‑of‑thought prompting while using far less inference compute.

By Zhaoliang Chen, Jie Fu
arXiv Computation and Language
3d ago

LatentHarness: Learning Latent Actions for Memory and Reasoning via Counterfactual Policy Distillation

LatentHarness unifies memory access and latent reasoning by treating them as sequential latent actions—THINK, RECALL, and EXIT—within a language model. It is trained via counterfactual policy distillation, which evaluates the impact of each action on the emitted token and learns when to recall evidence versus continue reasoning. On six long‑context reasoning benchmarks, a 1.4B‑parameter LatentHarness model outperforms the strongest baselines by 2.8% and 10.0% relative, while running 5.9× faster than the leading long‑context baseline.

By Xiaoqiang Wang, Suyuchen Wang, Bang Liu
arXiv Machine Learning
Sep 25

Thinking Leakage: A Causal Audit of NoThink Post-Training in Hybrid Reasoning Models

The paper investigates ‘thinking leakage’ in post‑training hybrid reasoning models that operate in NoThink mode. Using a causal mediation framework and bidirectional interventions, the authors show that the performance gains attributed to NoThink training largely stem from the model drifting toward its Think mode. Across three models and three training methods on competition math benchmarks, leakage ratios between 42% and 79% were observed, indicating that much of the accuracy improvement is due to re‑invoking existing Think behavior rather than genuine NoThink capability.

By Zehao Liu, Vasant G. Honavar
arXiv Computation and Language
Sep 1

Detecting Hidden Chain-of-Thought in Large Language Models with Linguistic, Behavioral, and Mechanistic Indicators

arXiv:2608.29956v1 Announce Type: new Abstract: Large language models often answer complex reasoning questions without revealing intermediate steps, raising whether they reason latently or complete p...

By Armaan Singh, Ryan Trinh Le, Jasmine Kaur, Abdullah Sultan, Edward Lue Chee Lip, Kiran Nijjer, Adnan Ahmed, Vasu Sharma