arXiv Machine Learning

What preferences can - and cannot - predict in multi-agent online learning

arXiv:2608. 13810v1 Announce Type: cross Abstract: We examine the interplay between ordinal, preference-based solution concepts in games and the long-run behavior of game dynamics, asking in particular to what extent the combinatorial data of a game -- its preference graph -- determine the outcomes of no-regret learning dynamics -- such as follow-the-regularized-leader (FTRL).

arXiv Machine Learning
Sep 22

Learning in Structured Stackelberg Games

arXiv:2504.09006v5 Announce Type: replace-cross Abstract: We initiate the study of structured Stackelberg games, a novel form of strategic interaction between a leader and a follower where contextual...

By Maria-Florina Balcan, Kiriaki Fragkia, Keegan Harris
arXiv AI
Sep 24

Evolutionary Stability Does Not Guarantee Learning Accessibility: A Multi-Agent Reinforcement Learning Perspective on Cooperation Emergence

The paper investigates whether evolutionary stability guarantees that learning agents can achieve cooperative outcomes in a multi‑agent setting. Using a three‑agent governance game, the authors compare the evolutionary basin of attraction with learning basins derived from independent Q‑learning, scaled Boltzmann exploration, and SA–EA BQL. They find that while the evolutionary basin covers the entire sampled grid, only ε‑greedy Q‑learning attains a substantial learning basin, whereas the other methods fail to sustain cooperation, highlighting a disconnect between population‑level stability and finite‑sample learning accessibility.

By Yijie Wang
arXiv Machine Learning
Jul 14

Paradoxes of Game Theoretic Equilibria and Price of Anarchy

arXiv:2607. 11752v1 Announce Type: cross Abstract: For decades, static solution concepts (Nash, Correlated, and Coarse Correlated Equilibria) and the Price of Anarchy (PoA) have formed the bedrock of algorithmic game theory, with no-regret learning proving fast convergence to such game-theoretic equilibria.

By Georgios Piliouras, Ian Gemp, Siqi Liu, Luke Marris