arXiv Machine Learning By Seonglae Cho, Zekun Wu, Kleyton Da Costa, Adriano Koshiyama

The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models

Read the original on arXiv Machine Learning →

arXiv:2602. 08159v2 Announce Type: replace Abstract: When a language model asserts that "the capital of Australia is Sydney," does it know this is wrong?

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.