Hugging Face Trending Papers

Best-of-Evidence: Best-of-N Selection under Partial Verification

Read the original on Hugging Face Trending Papers →

BoN improves model outputs by sampling several candidates and selecting one with a proxy score, but it assumes that complete candidates can be evaluated reliably. Many vision-language tasks instead provide only partial verification: a finding, span, value, region, or relation may be checkable even when no dependable whole-response verifier exists.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv Machine Learning
Jul 7

Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation

arXiv:2603. 10494v2 Announce Type: replace-cross Abstract: Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedback and can reward models that simply say less.

By Weixin Liu, Congning Ni, Qingyuan Song, Susannah L. Rose, Murat Kantarcioglu, Bradley A. Malin, Zhijun Yin
arXiv AI
Sep 18

The Missing Complement: State-Conditioned Minimal Sufficient Evidence for Coding Agents

The paper introduces State‑Conditioned Minimal Sufficient Evidence Recovery (SER), a method that, given a coding agent’s current state, reconstructs a compact set of evidence passages that collectively provide all facts needed for the agent’s next decision. Using the SERBench dataset of 500 held‑out states from 45 repositories, the authors show that their MSS‑Complement approach recovers a complete evidence set for 73.0 % of states with five items and 80.6 % with eight, outperforming baseline ranking methods. The study also demonstrates that this set‑level policy improves downstream performance on AMA‑Bench and highlights the importance of retrieving missing facts rather than merely re‑ranking similar passages.

By Zhexi Feng, Ruiyi Zhang, Yongbo Yang, Pengtao Xie
arXiv AI
Sep 24

Agentic Governance and Adversarial Verification for Policy-Constrained LLM Healthcare Appeal Generation

The paper introduces AGVF, a multi‑agent framework for generating medical‑necessity appeals that must adhere to payer policy and evidence constraints. AGVF treats appeal synthesis as a Constrained Markov Decision Process involving five agents—policy formalization, evidence retrieval, gap analysis, adversarial critique, and gated synthesis—and proves that refining the policy constraint graph monotonically reduces evidence deficiency. A deterministic citation‑grounding gate ensures no unsupported assertions enter the shared state, and validation on 1,000 synthetic cases shows zero citation violations and monotonic deficiency reduction.

By Harshil Lodhiya, Alex McManus, Reese Walker