Hugging Face Trending Papers

Can Trustless Agents Be Trusted? An Empirical Study of the ERC-8004 Decentralized AI Agent Ecosystem

Read the original on Hugging Face Trending Papers →

As autonomous AI agents increasingly transact across organizational boundaries, a fundamental trust challenge emerges: how can an agent assess whether an unknown counterpart is trustworthy? The ERC-8004 protocol addresses this challenge with the first permissionless trust layer for AI agent economies, built around three on-chain registries for Identity, Reputation, and Validation.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv AI
4d ago

Agentic Commerce Bench: Measuring Fraud Detection for Agents That Spend Money

The paper introduces the Agentic Commerce Bench (ACB), a benchmark for measuring fraud in AI agents that autonomously spend money. It presents a taxonomy of agentic commerce fraud, a dataset of twenty fraud classes derived from real production data, and an open‑source detector stack called gordonguard for auditing and replaying hostile counterparties. The study shows that current reasoning layers and security scanners perform poorly on many classes, highlighting the need for better detection mechanisms.

By Ankit Srivastava, Debjyoti Paul
arXiv AI
Sep 10

DART: A DAG-Based Reputation and Incentive Framework via Blockchain-Enabled Governance for Trustworthy LLM Multi-Agent Collaboration

The paper introduces DART, a Directed Acyclic Graph (DAG)-based framework that combines centralized orchestration with blockchain-enabled decentralized governance to manage reputation and incentives in large language model (LLM)-based multi-agent systems. DART dynamically allocates tasks based on agent capability, reputation, and workload, while continuously updating trust scores through post-execution evidence and smart contract accountability. Experimental results show that DART outperforms centralized baselines, achieving high task success rates, low retry rates, and effective containment of malicious agents.

By Manoj Kumala, Xinyun Liua, Ronghua Xu