arXiv AI By Bo Deng, Xinlei Zheng, Yi Wei, Kang Zhou, Chongyang Tao, Renzhao Liang, Xuanren Chen, Lifan Guo, Chi Zhang

DeFA: Dependency-Guided Failure Attribution for LLM Agents

Read the original on arXiv AI →

DeFA is a dependency-guided framework that attributes failures in large language model agents by constructing an event dependency graph and a failure propagation graph from protocol relations and semantic dependencies. It identifies violating events, traces their sources and effects, and determines the decisive error, responsible agent, and error category. The method supports long trajectories through segmentation and has shown superior accuracy on text, image, and video tasks, while its diagnostic feedback can improve agent performance on subsequent tasks.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 7

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

arXiv:2608. 06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging.

By Yunjia Qi, Zehua Yin, Xintong Shi, Hao Peng, Songyuanyi Lu, Yixian Liu, Richeng Xuan, Yuhong Liu, Zhichao Hu, Xiaozhi Wang, Lei Hou, Bin Xu, Juanzi Li
arXiv AI
Sep 2

EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems

EDGE is a framework that attributes multiple related errors in multi-agent large language model systems by constructing an error dependency graph from observed error events. It validates a reliable causal subset through counterfactual rollout and uses this inference graph to guide a two-stage LLM-as-judge detector for error attribution. Experiments on TRAIL and MAST demonstrate that EDGE improves category-level multi-error attribution across most models and settings, and that the graph aids explanation and repair analysis.

By Jun Hou, Priya Pitre, Yi Fang, Xuan Wang