arXiv AI

DART: A DAG-Based Reputation and Incentive Framework via Blockchain-Enabled Governance for Trustworthy LLM Multi-Agent Collaboration

The paper introduces DART, a Directed Acyclic Graph (DAG)-based framework that combines centralized orchestration with blockchain-enabled decentralized governance to manage reputation and incentives in large language model (LLM)-based multi-agent systems. DART dynamically allocates tasks based on agent capability, reputation, and workload, while continuously updating trust scores through post-execution evidence and smart contract accountability. Experimental results show that DART outperforms centralized baselines, achieving high task success rates, low retry rates, and effective containment of malicious agents.

Hugging Face Trending Papers
Jun 24

Can Trustless Agents Be Trusted? An Empirical Study of the ERC-8004 Decentralized AI Agent Ecosystem

As autonomous AI agents increasingly transact across organizational boundaries, a fundamental trust challenge emerges: how can an agent assess whether an unknown counterpart is trustworthy? The ERC-8004 protocol addresses this challenge with the first permissionless trust layer for AI agent economies, built around three on-chain registries for Identity, Reputation, and Validation.

arXiv AI
3d ago

PANDA: A Decentralized Architecture with Flexible Orchestration for Scalable, Fault-Tolerant Multi-Agent Systems

PANDA is a decentralized architecture for large-scale, fault-tolerant multi-agent systems that enables heterogeneous agents to discover each other's capabilities and self-organize into specialized teams for each task. It decouples collective communication from team communication, allowing agents to participate in multiple teams simultaneously and load-balance tasks across the collective. PANDA supports three planning and execution patterns—star, chain, and mesh—detects and recovers from infrastructure and orchestration failures, and uses a web-of-trust model for governance without a central bottleneck.

By Matthew D. Laws, Cristina Nita-Rotaru
arXiv AI
Sep 2

Delegation Without Trust: An Empirical Gap Analysis of Identity, Authorization, and Runtime Governance in Multi-Agent LLM Systems

The paper examines the security challenges of delegating authority to autonomous LLM agents that act on users’ behalf. It introduces a threat model with four adversaries and eight security requirements, demonstrates that current frameworks (LangGraph, CrewAI, AutoGen, MCP) fail to meet these standards, and presents an authorization broker that blocks all identified threats with minimal overhead. The broker is shown to resist numerous attacks and limits compromised sub‑agents to their delegated tasks, and its principles are implemented in VotalAI’s LLM Shield.

By Panduranga Sai Varma Dantuluri, Jyotirmoy Sundi
Hugging Face Trending Papers
Jul 9

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustworthiness. In decentralized energy markets, autonomous agents may improve market utility, but may also exploit invalid physical data, create artificial liquidity, and produce unstable governance decisions.

arXiv AI
Aug 5

WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks

arXiv:2608. 03499v1 Announce Type: new Abstract: Recent advances in persistent personal-agent frameworks are making human-centered agent networks realistic deployment targets: each user can be served by an AI agent that acts on the user's behalf, maintains state, and communicates with other agents through social and task relations.

By Prince Zizhuang Wang, Aojie Yuan, Haiyue Zhang, Xiyang Hu, Yue Zhao, Shuli Jiang