arXiv AI By Hankyeol Kim, Pilsung Kang

Same Answer, Different Confidence: Protocol Sensitivity in LLM Confidence Calibration

Read the original on arXiv AI →

arXiv:2605. 27752v3 Announce Type: replace Abstract: Is verbalized confidence better calibrated than token likelihood?

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.