arXiv AI By Maheep Chaudhary

In-Context Environments Induce Evaluation-Awareness in Language Models

Read the original on arXiv AI →

arXiv:2603. 03824v2 Announce Type: replace Abstract: Humans often become more self-aware under threat, yet can lose self-awareness when absorbed in a task; we hypothesize that language models exhibit environment-dependent \textit{evaluation awareness}.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.