arXiv Machine Learning By Govind Ramesh, Yao Dou, Wei Xu

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

Read the original on arXiv Machine Learning →

arXiv:2606. 05486v1 Announce Type: cross Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, while existing attribution methods are designed to explain observable outputs such as logits or generated tokens.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.