← Back to all news
OpenAI Blog April 10, 2017

Stochastic Neural Networks for hierarchical reinforcement learning

Read the original on OpenAI Blog →

The Flow has not summarised this story yet — read it at OpenAI Blog.

  • reinforcement-learning

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jun 4

From Ticks to Flows: Dynamics of Neural Reinforcement Learning in Continuous Environments

arXiv:2606. 04275v1 Announce Type: cross Abstract: We present a novel theoretical framework for deep reinforcement learning (RL) in continuous environments by modeling the problem as a continuous-time stochastic process, drawing on insights from stochastic control.

By Saket Tiwari, Tejas Kotwal, George Konidaris
reinforcement-learning
More like this →
Hugging Face Blog
May 4, 2022

An Introduction to Deep Reinforcement Learning

reinforcement-learning
More like this →
OpenAI Blog
Nov 9, 2016

RL²: Fast reinforcement learning via slow reinforcement learning

reinforcement-learning
More like this →
arXiv Machine Learning
2d ago

Action-Driven Processes for Continuous-Time Control

arXiv:2510.26672v3 Announce Type: replace-cross Abstract: At the heart of reinforcement learning are actions -- decisions made in response to observations of the environment. Actions are equally fund...

By Ruimin He, Shaowei Lin
reinforcement-learning
More like this →
arXiv Machine Learning
Jul 28

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning

arXiv:2607. 23726v1 Announce Type: cross Abstract: Exploration in sparse-reward long-horizon tasks poses significant challenges for reinforcement learning.

By Zahra Abdalla Elashaal, Afef Hfaiedh, Nahla Khraief, Issmail Ellabib, Giansalvo Cirrincione
reinforcement-learning
More like this →
arXiv AI
Aug 18

Adaptive Mixing of Policies from Searching and Policies from Learning

arXiv:2608. 15700v1 Announce Type: new Abstract: Background: Distillation of training targets generated thru search/planning has proven useful in reinforcement learning, but search can take exceedingly long.

By Gavin B. Rens
reinforcement-learningefficiency
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea