Hugging Face Trending Papers

Chronocooked: A Benchmark for Implicit Interval Timing in Reinforcement Learning Agents

This paper presents Chronocooked, a reinforcement learning (RL) benchmark suite for studying implicit interval timing in RL agents. Inspired by Overcooked, the suite comprises cooking scenarios that require temporal decision making.

arXiv AI
Jun 9

Engagement Process: Rethinking the Temporal Interface of Action and Observation

arXiv:2605. 11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over different time scales rather than align with fixed observation--action steps.

By Jialian Li, Yuchen Cao, Junhong Liu, Weiran Guo, Xutao Wang, Jiaming Song, Jiahao Zhang, Jie Chen
arXiv AI
Jul 14

Adaptive Reinforcement Learning for Unobservable Random Delays

arXiv:2506. 14411v2 Announce Type: replace-cross Abstract: In standard reinforcement learning (RL) settings, the interaction between the agent and the environment is typically modeled as a Markov decision process (MDP), which assumes that the agent observes the system state instantaneously, selects an action without delay, and executes it immediately.

By John Wikman, Alexandre Proutiere, David Broman