arXiv:2608.28977v1 Announce Type: new
Abstract: Several works have investigated the influence of graph topology on cooperation among artificial agents, while the majority of the literature has focuse...
By Seongho Son, Stephen Hailes, Mirco Musolesi
arXiv:1908.08773v3 Announce Type: replace
Abstract: In certain reinforcement learning (RL) scenarios there are adversaries trying to interfere with the underlying reward process for their own benefit...
By Victor Gallego, Roi Naveiro, David Rios Insua, David Gomez-Ullate Oteiza
arXiv:2606. 27909v1 Announce Type: cross Abstract: Theory-of-mind evaluations of large language models typically use dyadic social-deduction games, where every observable cue points to a single hidden side, so a model with strong language priors can score well without ever simulating opponents' incentives.
By Avni Mittal
arXiv:2511. 22226v2 Announce Type: replace Abstract: The standard theory of model-free reinforcement learning assumes that the environment dynamics are stationary and that agents are decoupled from their environment, such that policies are treated as being separate from the world they inhabit.
By Alexander Meulemans, Rajai Nasser, Maciej Wo{\l}czyk, Marissa A. Weis, Seijin Kobayashi, Blake Richards, Guillaume Lajoie, Angelika Steger, Marcus Hutter, James Manyika, Rif A. Saurous, Jo\~ao Sacramento, Blaise Ag\"uera y Arcas
arXiv:2609.38516v1 Announce Type: cross
Abstract: Large language models (LLMs) can now improve themselves by revising the instructions they follow, and LLM agents are increasingly orchestrated to wor...
By Kunal Jha, Max Kleiman-Weiner, Natasha Jaques