arXiv AI

Moloch's Bargain: Emergent Misalignment When LLMs Compete for Audiences

arXiv AI
Aug 25

CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulation

arXiv:2604.09746v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly deployed as autonomous agents, understanding how strategic behavior emerges in multi-agent e...

By Aarush Sinha, Arion Das, Soumyadeep Nag, Charan Karnati, Shravani Nag, Chandra Vadhan Raj, Aman Chadha, Vinija Jain, Suranjana Trivedy, Amitava Das
Hugging Face Trending Papers
Aug 12

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs

Persuasion is a core dynamic of natural language communication, shaping how large language models (LLMs) update beliefs, resolve disagreements, and reach decisions. As LLMs increasingly debate, advise, and think collaboratively with humans and each other, resistance to harmful persuasion becomes a core requirement for reliable behavior.

arXiv AI
Aug 18

Emergent Misaligned Communication in Long-Horizon Multi-Agent LLM Commerce

arXiv:2608. 14825v1 Announce Type: cross Abstract: Frontier LLM agents increasingly transact on behalf of separate principals, often using natural language rather than structured APIs.

By Zeyuan Li (Massachusetts Institute of Technology), Lukas Petersson (Andon Labs), Alessandro Acquisti (Massachusetts Institute of Technology), Michiel A. Bakker (Massachusetts Institute of Technology)
arXiv Machine Learning
Sep 25

Why Does Misinformation Propagate Faster? An Algorithmic Perspective on X

The paper investigates why misinformation spreads more quickly on engagement‑based platforms by dissecting the recommendation algorithm of X. It identifies an engagement fungibility mechanism that rewards instant reactions (likes, retweets) over thoughtful engagement (replies, quotes), allowing misinformation—which tends to attract instant reactions—to receive more recommendations. The authors validate this mechanism through a simulation on the USC X 2024 election corpus, showing that adjusting metric weights has little effect, while requiring thoughtful engagement before amplification can significantly reduce the credibility exposure gap without harming mainstream content or engagement.

By Pan Li, Shuang Gao
arXiv Computation and Language
Sep 22

The Role of AI in Online Reviews

arXiv:2609.22198v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) creates new opportunities for strategic content generation on online platforms, including potentiall...

By Valeria Lerman, Oren Rigbi, Yaniv Dover
arXiv Computation and Language
Aug 27

Belief Cascades Drive Persuasion in LLM Agent Networks

The paper introduces a controlled testbed to study how goal‑directed persuaders shift stances in networks of large language model agents, using real‑world ego‑network topologies. Experiments across four LLM backbones, five graph structures, and 55 policy statements show that persuasion dynamics depend on topology, competition, topic, and model prior. The study finds that direct exposure predicts stance change, peer relays have measurable influence, and that post‑text analysis alone misses important movement, highlighting the need to evaluate multi‑agent persuasion through trajectory‑level processes, belief probes, exposure provenance, and action logs.

By Haoyi Qiu, Genglin Liu, Pranav Narayanan Venkit, Kung-Hsiang Huang, Saadia Gabriel, Chien-Sheng Wu, Nanyun Peng