arXiv AI By Xian Sun, Wei Chow, Yingshuo Wang, Junhao Liu, Wei Gao, Qing Wu, Lingdong Kong

Learning When to Trust via Selective Context Preference Optimization

Read the original on arXiv AI →

arXiv:2608. 06377v1 Announce Type: cross Abstract: Language models increasingly condition their answers on external signals, and a single misleading one can turn a correct answer wrong.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.