arXiv:2608. 16578v1 Announce Type: new Abstract: AI agents increasingly operate as part of interacting systems rather than in isolation.
By Batu El, Jinhee Paeng, Fatih Dinc, Shiye Su, Mete Erdogan, Aneesh Pappu, Haotian Ye, Wanjia Zhao, Surya Ganguli, James Zou
The Flag Game is a toy model designed to study how AI agents form collective beliefs. In the game, each agent sees only a private crop of a hidden country flag and can share beliefs with peers, leading to complex phenomena such as non‑monotonic performance scaling, accuracy gains from social awareness, and polarization that degrades performance at large population sizes. The authors introduce social circuit attribution to identify key agents and views, and develop a statistical mechanical theory to explain collective belief collapse and polarization in larger populations.
By Elizabeth Pavlova, Hidenori Tanaka
Emergent coordinated behaviors of AI agents are starting to present critical safety risks. A key phenomenon driving these behaviors is the rapid formation and spread of beliefs about the world, and me...
The paper introduces Autopoietic Game Theory, a computational model where social interactions, replication mechanisms, and computational costs co-evolve within a substrate of randomly initialized Z80 machine code programs. By embedding a social dilemma directly into the physics of computation, the authors demonstrate that scarcity of resources can make defection self-limiting, leading to the emergence of self-replicating, cooperative strategies. Empirical results show evolved programs suppress stealing, and spatial assortment enhances structural complexity and task performance, while the framework can also incorporate exogenous pressures such as math tasks tied to computation budgets.
By Kunal Jha, Francesco Cicala, Blaise Ag\"uera y Arcas, Blake Aaron Richards, Natasha Jaques, Max Kleiman-Weiner, Eyvind Niklasson
arXiv:2608.28046v1 Announce Type: cross
Abstract: Collective behaviour in living systems is usually modelled as the outcome of a \emph{direct} social drive: agents are rewarded, or hard-wired, to ali...
By Gorka Mu\~noz-Gil, Andrea L\'opez-Incera, Vide Ramsten, Giovanni Volpe, Thomas M\"uller, Hans J. Briegel
arXiv:2605. 30169v2 Announce Type: replace-cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to decide whether to trust an unfamiliar agent in the wild and delegate to it?
By Botao Amber Hu, Helena Rong, Max Van Kleek