arXiv AI

Think-Before-Speak: From Internal Evaluation to Public Expression in Multi-Agent Social Simulation

arXiv:2606. 03137v1 Announce Type: new Abstract: LLM-based multi-agent simulation offers a promising way to study social interaction, deliberation, and collective opinion dynamics.

arXiv AI
Aug 28

Can You Say This for Me? Speaking Up by Proxy in Co-Located Discussion

SecondVoice is a mixed‑reality system that lets participants speak up in co‑located discussions via an embodied virtual proxy, separating the content of a point from the speaker’s identity. Users input their intent through a structured process, and the system reformulates and vocalizes it into the conversation. In a study with 16 participants, the proxy channel led to more points being voiced than an anonymous text board and sparked multi‑turn engagement that the text board did not elicit.

By Yue Shen, Rehema Abulikemu, Ryan P. McMahan, Yan Chen
arXiv AI
Sep 10

From Simulated Citizens to Simulated Deliberation: Challenges in Representation and Interaction

The paper investigates whether large language model (LLM) agents can simulate public deliberation by reflecting population opinion patterns and producing interaction-driven opinion change. Using census‑grounded Korean personas debating real policy questions, the study finds that persona agents fail to reliably reproduce population opinion patterns, often concentrating responses and reversing demographic differences. While deliberations generate reasoned, reciprocal arguments and some stance movement, much of this change occurs without peer exchange, and anchoring agents to population‑informed starting positions suppresses updating, indicating that population representation, argument generation, and interaction‑driven opinion change do not necessarily align.

By Chaemin Jang, Junsik Min, Jaewoo Choi, Donggyu Lee, Haiin Lee, Junyoung Park, Namhee Kim, Hyunwoo Kim, Jungwon Kim, Juho Kim, Nuri Kim, Jihee Kim
arXiv Computation and Language
Aug 31

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations?

The paper investigates whether large language models (LLMs) replicate socio‑cognitive effects of power asymmetry observed in human communication. By assigning high or low status personas to LLMs in simulated multi‑turn dialogues across diverse professions, the study measures language coordination, pronoun usage, persuasion success, and compliance with unsafe requests. Results indicate that LLMs exhibit key power‑related socio‑cognitive behaviors, though with nuances and variability, linking these simulated interactions to both desirable and unsafe outcomes.

By Anvesh Rao Vijjini, Sagar Manjunath, Snigdha Chaturvedi
arXiv AI
6d ago

Thinking Less to Simulate Better: Intuitive Prompting Improves LLM Agents Simulating Individual Social Media Reactions, Including Unfamiliar Content

The study evaluates how well language‑model agents can simulate individual social media reactions by comparing predictions under different prompt conditions. Eight Serbian participants’ reactions to 68 posts were recorded, and four language models were asked to predict these reactions using prompts that varied in profile content and instruction style. The results show that prompts emphasizing attitudinal content and intuitive, immediate responses yield the highest fidelity, outperforming demographic backstories and a crowd baseline, and suggesting that such agents could act as general‑purpose simulated users.

By Ljubisa Bojic, Tijana Stanic, Joerg Matthes, Agariadne Dwinggo Samala, Bojana Dinic, Jue Wang
arXiv Computation and Language
Sep 11

The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies

The study audits 576 LLM-based social simulations from 350 papers using the PIMMUR framework, which evaluates agent profile, interaction, memory, minimal control, unawareness, and realism. Results show that PIMMUR principles are met more often than minimal control, unawareness, and realism, with frontier LLMs correctly identifying the underlying experiment in 65.2% of cases and half of prompts pre‑determining outcomes. Reproducing five experiments revealed that many reported collective phenomena disappear or reverse when PIMMUR principles are enforced, suggesting that apparent emergent behaviors may be methodological artifacts rather than genuine social dynamics.

By Jiaxu Zhou, Jen-tse Huang, Xuhui Zhou, Man Ho Lam, Xintao Wang, Hao Zhu, Wenxuan Wang, Maarten Sap