arXiv:2603.02935v2 Announce Type: replace
Abstract: Offline meta-reinforcement learning seeks to learn a policy that generalizes to new related tasks online. Context-based methods infer a task repres...
By Mohammadreza Nakheai, Aidan Scannell, Kevin Luck, Joni Pajarinen
arXiv:2608. 05989v1 Announce Type: new Abstract: Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL).
By Xinwei Liu, Junyuan Liang, Jianting Zhang, Wuhui Chen
arXiv:2605. 26012v2 Announce Type: replace-cross Abstract: Deep reinforcement learning (RL) agents commonly rely on high-dimensional neural representations, despite growing evidence that task-relevant value and policy structure may be intrinsically low-dimensional.
By Aleksandar Todorov, Matthia Sabatelli
arXiv:2501. 14622v5 Announce Type: replace Abstract: Learning efficient representations for decision-making policies is a challenge in imitation learning (IL).
By Aleksandar Vujinovic, Aleksandar Kovacevic
The paper introduces a neurosymbolic world model that separates observation reconstruction from reward prediction, enabling the model to adapt zero‑shot to new reward functions defined over a shared symbolic state space. This approach addresses the task‑dependency of traditional neural world models, which learn latent representations tied to specific training tasks. Experiments show that the neurosymbolic formulation generalises more strongly than purely neural methods.
By Isidoro Tamassia, Lennert De Smet, Giuseppe Marra
Joint-embedding predictive architectures (JEPAs) learn latent dynamics for planning and avoid representation collapse by matching features to maximum-entropy distributions such as isotropic Gaussians,...