arXiv:2609.12444v1 Announce Type: cross
Abstract: Simulated societies of large language model agents are used to study online polarization, and separately to study collective intelligence, but the tw...
By Raad Bin Tareaf
arXiv:2604.11312v3 Announce Type: replace-cross
Abstract: Large Language Models are increasingly deployed as interacting agents in settings such as online platforms, recommendation systems, and multi...
By Erica Cau, Andrea Failla, Giulio Rossetti
The Flag Game is a toy model designed to study how AI agents form collective beliefs. In the game, each agent sees only a private crop of a hidden country flag and can share beliefs with peers, leading to complex phenomena such as non‑monotonic performance scaling, accuracy gains from social awareness, and polarization that degrades performance at large population sizes. The authors introduce social circuit attribution to identify key agents and views, and develop a statistical mechanical theory to explain collective belief collapse and polarization in larger populations.
By Elizabeth Pavlova, Hidenori Tanaka
The paper investigates how a minority of biased agents in a multi‑agent system of large language models (LLMs) can amplify bias through textual interactions. Even a small percentage of persistently extreme agents causes significant opinion shifts among the non‑biased agents, with the effect occurring faster in the Llama 3.2 model than in a classical Friedkin‑Johnsen model. Semantic analysis shows that rhetorical consistency rises with biased exposure and that non‑biased agents adopt the biased vocabulary even when their numerical opinions change only modestly.
By Omran Berjawi, Giuseppe Fenza, Rida Khatoun
Emergent coordinated behaviors of AI agents are starting to present critical safety risks. A key phenomenon driving these behaviors is the rapid formation and spread of beliefs about the world, and me...
arXiv:2606. 19494v1 Announce Type: new Abstract: Multi-agent LLM deliberation, where agents exchange and revise answers over several rounds, is increasingly used to improve reasoning and accuracy, yet how and why it works is rarely modelled.
By Apurba Pokharel, Ram Dantu