arXiv AI By Yuyang Dai, Yuxia Wang

Rescaling Confidence: What Scale Design Reveals About LLM Metacognition

Read the original on arXiv AI →

arXiv:2603. 09309v2 Announce Type: replace Abstract: Verbalized confidence, in which LLMs report a numerical certainty score, is widely used to estimate uncertainty in black-box settings, yet the confidence scale itself (typically 0--100) is rarely examined.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.

arXiv AI
Jul 9

Measuring the metacognition of AI

arXiv:2603. 29693v3 Announce Type: replace Abstract: A robust decision-making process must take into account uncertainty, especially when the choice involves inherent risks.

By Richard Servajean, Philippe Servajean