arXiv AI By Enric Boix-Adsera, Benedict Tessler

Model Hypnosis: Strong control of AI via additive subliminal effects

Read the original on arXiv AI →

arXiv:2608. 16834v1 Announce Type: cross Abstract: We demonstrate that AI models are broadly susceptible to a phenomenon we call model hypnosis, in which individually weak and seemingly irrelevant cues in the prompt can be systematically combined to strongly control model behavior.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.