← Back to all news
arXiv Machine Learning October 2, 2026 By Sathya Kamesh Bhethanabhotla, Efstratios Gavves, Andr\'e Biedenkapp

Reinforcement Learning with Complex (valued) Memories

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

  • agents
  • reinforcement-learning

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Sep 22

Streaming Deep Reinforcement Learning Finally Works

arXiv:2410.14606v3 Announce Type: replace Abstract: Learning from a stream of experience as it arrives, also known as streaming learning, is a core part of natural learning. However, reliable streami...

By Mohamed Elsayed, Elena Sorina Lupu, Gautham Vasan, A. Rupam Mahmood
reinforcement-learningroboticsbenchmarks
More like this →
arXiv Machine Learning
Jun 5

Unraveling the Hidden Dynamical Structure in Recurrent Neural Policies

arXiv:2602. 01196v2 Announce Type: replace Abstract: Recurrent neural policies are widely used in partially observable control and meta-RL tasks.

By Jin Li, Yue Wu, Mengsha Huang, Yuhao Sun, Hao He, Xianyuan Zhan
reinforcement-learning
More like this →
arXiv Machine Learning
Jun 4

Episodic Memory Temporal Consistency for Cooperative Multi-Agent Reinforcement Learning

arXiv:2606. 04492v1 Announce Type: new Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) frequently suffers from severe reward sparsity and exploration bottlenecks.

By Zicheng Zhao, Yu Lan, Chengzhengxu Li, Zhaohan Zhang, Xiaoming Liu
agentsreinforcement-learningefficiencybenchmarks
More like this →
arXiv Machine Learning
Jul 8

Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

arXiv:2605. 24709v2 Announce Type: replace Abstract: Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process data incrementally, i.

By Noah Farr, Aryaman Reddi, Carlo D'Eramo, Jan Peters
agentsreinforcement-learning
More like this →
arXiv Machine Learning
2d ago

Continual Learning of Dynamical Systems in Recurrent Neural Networks through Recyclable Unit Gating

arXiv:2609.38356v1 Announce Type: new Abstract: Dynamical Systems Reconstruction (DSR) aims to infer models from observed time series that reproduce a system's qualitative long-term behavior. Continu...

By Sima Hashemi, Daniel Durstewitz, Georgia Koppe
agentsbenchmarks
More like this →
arXiv Machine Learning
Sep 21

REFINEPPO: Learning Continuous Control Policies by Iterative Action Refinement

arXiv:2609.21108v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) has achieved strong performance across a wide range of continuous-control problems. These continuous-control policies...

By Sachini Weerasekara, Sagar Kamarthi, Jacqueline Isaacs
agentsreinforcement-learningbenchmarks
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea