arXiv:2606. 26463v1 Announce Type: new Abstract: Deliberating takes time.
By Aneesh Muppidi, Firas Darwish, Dylan Cope, Jo\~ao F. Henriques, Jakob Nicolaus Foerster
Deliberating takes time. In real-time settings, that time is not free.
arXiv:2607. 05064v1 Announce Type: new Abstract: Reinforcement learning in real world environments often suffers from severe performance degradation due to delayed feedback.
By Junqi Tu, Zejiao Liu, Fangfei Li, Yang Tang
arXiv:2603. 03480v2 Announce Type: replace Abstract: We study reinforcement learning with delayed state observation, where the agent observes the current state after some random number of time steps.
By Harin Lee, Kevin Jamieson
arXiv:2608. 11511v1 Announce Type: cross Abstract: In sequential decision making, an agent typically observes its environment and acts at every timestep.
By Christopher Watson, Arjun Krishna, Dinesh Jayaraman, Rajeev Alur
arXiv:2607. 16421v1 Announce Type: new Abstract: It has long been recognized that humans have the ability to switch between fast, reactive decision-making and slower, deliberative planning.
By Adam Labiosa, Josiah P. Hanna