arXiv Machine Learning By Han-yu Wang

Priors Persist Through Suppression: A Stroop Paradigm for Lexical Override

Read the original on arXiv Machine Learning →

arXiv:2606. 07555v1 Announce Type: cross Abstract: Glossaries, technical specifications, and system prompts routinely ask language models to use familiar words in unfamiliar ways.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.