arXiv AI By Saeed Ahmadnia, Cornelia Caragea

ReHoPER: Receding-Horizon Planning for Enhanced Reasoning

Read the original on arXiv AI →

ReHoPER is an inference‑only, zero‑shot method that enhances large language models’ reasoning by generating and answering intermediate questions along multiple paths before producing a final answer. It plans a horizon of candidate intermediate questions, selects one to answer, and replans based on the updated history. The approach is task‑agnostic, using generic instructions across datasets and models without labeled data or task‑specific prompt design, and it outperforms strong baselines on several datasets, notably achieving the largest gains on the new iLLC benchmark for compositional reasoning.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 9

MixReasoning: Switching Modes to Think

arXiv:2510. 06052v2 Announce Type: replace Abstract: Reasoning models enhance performance by tackling problems in a step-by-step manner, decomposing them into sub-problems and exploring long chains of thought before producing an answer.

By Haiquan Lu, Gongfan Fang, Xinyin Ma, Qi Li, Xinchao Wang
arXiv AI
Jul 14

PiCSAR: Probabilistic Confidence Selection And Ranking for Reasoning Chains

arXiv:2508. 21787v3 Announce Type: replace-cross Abstract: Best-of-n sampling improves the accuracy of large language models (LLMs) and large reasoning models (LRMs) by generating multiple candidate solutions and selecting the one with the highest reward.

By Joshua Ong Jun Leang, Zheng Zhao, Aryo Pradipta Gema, Sohee Yang, Wai-Chung Kwan, Xuanli He, Wenda Li, Pasquale Minervini, Eleonora Giunchiglia, Shay B. Cohen