arXiv AI By Youting Wang, Yuan Tang, Yitian Qian, Chen Zhao

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

Read the original on arXiv AI →

arXiv:2606. 07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking external tools.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.