← Back to all news
arXiv Computation and Language October 5, 2026 By Adithya Bhaskar, Jeffrey Cheng, Danqi Chen

Language Models that Play Chess and Explain Their Moves

Read the original on arXiv Computation and Language →

The Flow has not summarised this story yet — read it at arXiv Computation and Language.

  • llms
  • robotics
  • efficiency

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jun 24

Grounded Chess Reasoning in Language Models via Master Distillation

arXiv:2603. 20510v2 Announce Type: replace Abstract: Language models often lack grounded reasoning capabilities in specialized domains where training data is scarce but bespoke systems excel.

By Zhenwei Tang, Qianfeng Wen, Seth Grief-Albert, Yahya Elgabra, Blair Yang, Honghua Dong, Ashton Anderson
llmsreinforcement-learningfine-tuningefficiency
More like this →
arXiv AI
Jul 20

Understanding Reasoning from Pretraining to Post-Training

arXiv:2607. 16097v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is largely studied in isolation from the pretraining that precedes it.

By Jingyan Shen, Ang Li, Salman Rahman, Yifan Sun, Micah Goldblum, Matus Telgarsky, Pavel Izmailov
llmsreinforcement-learningfine-tuning
More like this →
arXiv AI
Aug 7

Correct Answers from Sound Reasoning: Verifiable Process Supervision for Language Models

arXiv:2605. 12519v2 Announce Type: replace-cross Abstract: Training language models to produce both correct answers and sound reasoning remains an open challenge.

By Kyuyoung Kim, Kevin Wang, Yunfei Xie, Peiyang Xu, Peiyao Sheng, Chen Wei, Zhangyang Wang, Jinwoo Shin, Pramod Viswanath, Sewoong Oh
llmsreinforcement-learningfine-tuning
More like this →
arXiv Machine Learning
Jul 27

LeAct: Learning to Reason from Expert Actions

arXiv:2607. 21856v1 Announce Type: new Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs.

By Ziran Yang, Chengshuai Shi, Raj Ghugare, Benjamin Eysenbach, Karthik Narasimhan, Chi Jin
llmsroboticsbenchmarks
More like this →
arXiv AI
Sep 2

Exploring Collaboration between a language and a non-language agent

arXiv:2609.00474v1 Announce Type: cross Abstract: LLMs are increasingly deployed as orchestrators that coordinate specialized subagents to solve complex tasks through natural language. However, in ma...

By Harini S I, Somesh Singh, Yaman K Singla, Rajiv Ratn Shah, David Doermann, Balaji Krishnamurthy
llmsagentsroboticsbenchmarks
More like this →
arXiv Computation and Language
Sep 22

Do Chess Explanations Reflect Model Decisions? Behavioral and Token-Level Tests of LLM Reasoning Faithfulness

arXiv:2609.22245v1 Announce Type: new Abstract: Large language models can produce fluent explanations for chess moves, but plausible language does not necessarily reflect the reasoning behind a decis...

By Angelina Parfenova
llms
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea