arXiv AI By Noah Y. Siegel, Nicolas Heess, Maria Perez-Ortiz, Oana-Maria Camburu

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

Read the original on arXiv AI →

arXiv:2503. 13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.