Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
Read the original on arXiv AI →The article reports that in July 2025, 18 arXiv manuscripts contained hidden instructions designed to manipulate AI‑assisted peer review, such as covert commands to give only positive reviews. These prompts were concealed using white text and microscopic fonts, and the authors’ reactions ranged from withdrawal to defending the practice as a test of reviewer misuse of large language models. The study identifies four types of hidden prompts, critiques the ineffectiveness of honeypot defenses, and highlights inconsistent publisher policies while calling for controlled AI integration and harmonized guidelines in academic evaluation.
Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.