arXiv AI

CrackedPDFs: A Controlled Benchmark for Hidden Prompt Injection in PDFs

arXiv:2607. 19396v1 Announce Type: new Abstract: Document-based LLM systems often flatten a PDF before guardrails inspect it.

arXiv Machine Learning
Jul 30

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

arXiv:2607. 26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning: maliciously injected passages that manipulate retrieved evidence.

By Pushkal Kumar, Tucker Nielson, Tanish Kolhe, Shubham Zala, Vincent Li