The paper introduces CAMS, a Claim‑Anchored Multi‑Document Summarization framework that decomposes source documents into atomic claims, resolves provenance deterministically from verbatim quotes to token spans, clusters equivalent claims across documents, and rewrites summaries so each sentence ends with claim identifiers linking back to source spans. CAMS separates provenance (an invariant for each emitted sentence) from faithfulness (an objective encouraged by selection, rewriting, and verification). Evaluations on MultiNews, DiverseSumm, and zero‑shot WCEP show that CAMS matches strong baselines in summary quality while improving faithfulness and citation precision, raising attribution accuracy from 38% to 64% and reducing human verification time per claim by 3.4×.
By Shuo Guan
arXiv:2608. 03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review.
By Peijia Guo, Wenxuan Xie, ZiGuang Li, Ming Li
The paper investigates how the granularity of citations—sentence, paragraph, or document level—affects the performance of large language models in attributed generation tasks. Across models ranging from 8B to 120B parameters, enforcing fine‑grained, sentence‑level citations consistently reduces performance, with median losses of 40% and up to 338% on specific tasks, while overall answer correctness remains largely unchanged. The study finds that attribution quality peaks at intermediate, paragraph‑level granularity, suggesting that overly fine citations break semantic dependencies and overly coarse ones add noise, and that the optimal granularity depends on model scale and the amount of evidence required.
whyItMatters":"The findings reveal that the conventional preference for fine‑grained citations can actually harm model performance, indicating that attribution standards should be tailored to the model’s semantic scope rather than fixed by convention."
By Hexuan Wang, Jingyu Zhang, Benjamin Van Durme, Daniel Khashabi
arXiv:2606. 26449v1 Announce Type: cross Abstract: Retrieval-augmented systems routinely present citations alongside generated answers, yet a citation does not confirm that the corresponding source meaningfully shaped the output.
By Mohammad Faizan, Dalal Alharthi
arXiv:2609.00241v1 Announce Type: new
Abstract: Long documents often distribute important information across extensive narrative passages and multiple tables, making faithful summarization particular...
By Meng Zhou, Wenhao You, Wei Yuan
arXiv:2607. 23804v1 Announce Type: cross Abstract: Context attribution methods for large language models (LLMs) identify which input context contributes to the model response.
By Quoc-Huy Trinh, Lin Zhu, Sebastian Szyller