arXiv Machine Learning By Teng-Ruei Chen

Resample or Reroute? Recoverable Stopping Debt Without Identified Action Selection

Read the original on arXiv Machine Learning →

The paper investigates how to handle a large‑language‑model’s weak verifier that accepts a response, followed by a second call that may resample or reroute. It introduces three decision gates—recoverable stopping debt, two‑sided FIT action support, and held‑out value from an outcome‑blind selector—to determine action selection, which is treated as an identification problem. Experiments on MBPP+, LiveCodeBench, and BigCodeBench show that while stopping debt exists, the current evidence does not identify when to resample versus reroute.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Aug 14

Dead text or binding clause? Measuring and restoring constraint influence in black-box LLM dialogues

arXiv:2608. 12599v1 Announce Type: new Abstract: Multi-turn dialogues let users revoke constraints as easily as impose them, but revocation does not reliably take effect: models keep enacting withdrawn requirements (occasionally beneath comments asserting their removal), a failure we call \emph{behavioral relapse}, or revocation inertia.

By Haoyuan Zhu
arXiv AI
Jun 6

Answer Presence Drives RAG Rewriting Gains

arXiv:2606. 05633v1 Announce Type: new Abstract: Retrieval-augmented QA pipelines often route retrieved passages through an LLM \emph{rewriter} before a smaller reader, lifting F1 by tens of points on multi-hop benchmarks; this gain is typically credited to improved evidence quality.

By Yuejie Li, Yueying Hua, Ke Yang, Li Zhang, Yueping He, Yueping He, Ruiqi Li, Bolin Chen, Tao Wang, Bowen Li, Chengjun Mao