arXiv AI By Maosen Zhang, Jianshuo Dong, Boting Lu, Wenyue Li, Xiaoping Zhang, Tianwei Zhang, Jie Zhang, Han Qiu

The Model's Tell: Measuring Context-Leakage Attack Signals with Behavior Gauges

Read the original on arXiv AI →

arXiv:2608. 17829v1 Announce Type: cross Abstract: LLMs increasingly rely on external contexts, such as pre-defined system prompts or retrieved documents, to improve generation quality.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.