arXiv AI
Jul 13

Heterogeneous Information-Bottleneck Coordination Graphs for Multi-Agent Reinforcement Learning

arXiv:2605. 17393v2 Announce Type: replace Abstract: Coordination graphs are a central abstraction in cooperative multi-agent reinforcement learning (MARL), yet existing sparse-graph learners lack a theoretically grounded mechanism to decide which edges should exist and how much information each edge should carry.

By Wei Duan, Junyu Xuan, En Yu, Xiaoyu Yang, Jie Lu
OpenAI Blog
Sep 14, 2017

Learning to model other minds

We’re releasing an algorithm which accounts for the fact that other agents are learning too, and discovers self-interested yet collaborative strategies like tit-for-tat in the iterated prisoner’s dilemma. This algorithm, Learning with Opponent-Learning Awareness (LOLA), is a small step towards agents that model other minds.