← Back to all news
arXiv Machine Learning August 26, 2026 By Xiewei Ni, Ruofeng Mei, Xiangyu Xu

CoDrift: Compositional Drifting for Offline Reinforcement Learning

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

  • reinforcement-learning
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Aug 24

Decoupling Policy Extraction for Offline Reinforcement Learning

arXiv:2608.20909v1 Announce Type: new Abstract: Offline RL methods commonly jointly train the actor and critic, where the critic is used to guide the actor toward higher-value actions. This coupled l...

By Xuyao Lin, Yixiang Shan, Jinru Duan, Tao Yang, Xinyu Zhao, Runyu Lei, Yiming Zhao, Jiaxin Fan, Zongbao Feng, Peng Jia
ragreinforcement-learning
More like this →
arXiv Machine Learning
3d ago

PathBridger: Subgoal Bridges for Offline Goal-Conditioned Reinforcement Learning

arXiv:2608.29061v1 Announce Type: new Abstract: Offline goal-conditioned reinforcement learning (GCRL) aims to learn policies for reaching diverse goals entirely from fixed trajectory data. Long-hori...

By Soohyun Choi, Seonvin Cho, Songnam Hong
reinforcement-learningrobotics
More like this →
arXiv Machine Learning
Jul 28

Hierarchical Reinforcement Learning with Optimal Level Synchronization Based on Flow-Based Deep Generative Model

arXiv:2107. 08183v2 Announce Type: replace Abstract: High-dimensional state and action spaces combined with sparse reward structures in reinforcement learning (RL) environments typically require advanced control architectures.

By JaeYoon Kim, Junyu Xuan, Christy Liang, Farookh Hussain
reinforcement-learningbenchmarks
More like this →
arXiv Machine Learning
Jun 16

Efficient Reinforcement Learning by Guiding World Models with Non-Curated Data

arXiv:2502. 19544v3 Announce Type: replace Abstract: Leveraging offline data is a promising way to improve the sample efficiency of online reinforcement learning (RL).

By Yi Zhao, Aidan Scannell, Wenshuai Zhao, Yuxin Hou, Tianyu Cui, Le Chen, Dieter B\"uchler, Arno Solin, Juho Kannala, Joni Pajarinen
reinforcement-learningroboticsfine-tuning
More like this →
arXiv AI
Jul 17

RAD: Retrieval High-quality Demonstrations to Enhance Decision-making

arXiv:2507. 15356v2 Announce Type: replace Abstract: Offline reinforcement learning (RL) learns policies from fixed datasets, thereby avoiding costly or unsafe environment interactions.

By Lu Guo, Yixiang Shan, Zhengbang Zhu, Qifan Liang, Lichang Song, Ting Long, Weinan Zhang, Yi Chang
agentsreinforcement-learningbenchmarks
More like this →
arXiv Machine Learning
1d ago

Recursive Value Learning for Long-Horizon Offline Goal-Conditioned RL

arXiv:2609.02237v1 Announce Type: new Abstract: Scaling offline goal-conditioned reinforcement learning (GCRL) to long-horizon tasks is difficult because (1) long-range value learning depends on shor...

By Hyeonseong Jeon, Youngwoon Lee
reinforcement-learning
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea