arXiv AI By Yunbo Tang, Chengyi Yang, Shiyu Liu, Zhishang Xiang, Zerui Chen, Qinggang Zhang, Jinsong Su

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

Read the original on arXiv AI →

arXiv:2605. 29796v3 Announce Type: replace Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
4d ago

Traverse: Learning When to Remember, Reset, and Redirect for Long-Horizon Web Search

The paper introduces Traverse, an autonomous web‑search agent that manages its search process through three states—Rubric, Answer, and Verify—while using a Seal Memory tool for active context management. Reinforcement learning is employed to train the agent, but a training instability called Seal Collapse is mitigated by training only the final segment after context management. The resulting 35B model achieves state‑of‑the‑art performance on BrowseComp and related benchmarks, outperforming comparable open‑source systems.

By Jingyuan Ma, Lynx Aster, He Zhang, Siyao Song, Weijie Yuan, Zhe Zhang, Kai Jia, Zhifang Sui
arXiv Computation and Language
Sep 2

AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning

AdaSearch introduces a two‑stage reinforcement learning framework that separates problem solving from the decision to search in large language models. By using an F1‑based decision metric, it explicitly evaluates when external search is needed, reducing unnecessary search calls while maintaining high question‑answering performance. Experiments show that AdaSearch improves search‑decision quality with only a minor impact on accuracy compared to always‑search strategies.

By Tzu-Han Lin, Wei-Lin Chen, Chen-An Li, Hung-yi Lee, Yun-Nung Chen, Yu Meng