← Back to all news
OpenAI Blog August 13, 2018

Large-scale study of curiosity-driven learning

Read the original on OpenAI Blog →

The Flow has not summarised this story yet — read it at OpenAI Blog.

Related stories

arXiv AI
Jun 19

Can In-Context Learning Support Intrinsic Curiosity?

arXiv:2606. 19476v1 Announce Type: cross Abstract: Effective machine learning depends not only on how we model data, but also on what data we choose to collect.

By Eric Elmoznino, Sangnie Bhardwaj, Johannes von Oswald, Rajai Nasser, Blaise Ag\"uera y Arcas, Jo\~ao Sacramento, Rif A. Saurous, Guillaume Lajoie
llmsagentsreinforcement-learningsafety
More like this →
OpenAI Blog
Mar 3, 2018

Some considerations on learning to explore via meta-reinforcement learning

reinforcement-learning
More like this →
Hugging Face Trending Papers
Jun 17

Can In-Context Learning Support Intrinsic Curiosity?

Effective machine learning depends not only on how we model data, but also on what data we choose to collect. While large sequence models have revolutionized data modeling, the problem of automated data selection, or "intrinsic curiosity", remains a significant challenge.

llmsagentsreinforcement-learningsafety
More like this →
OpenAI Blog
Oct 31, 2018

Reinforcement learning with prediction-based rewards

We’ve developed Random Network Distillation (RND), a prediction-based method for encouraging reinforcement learning agents to explore their environments through curiosity, which for the first time exceeds average human performance on Montezuma’s Revenge.

agentsreinforcement-learningefficiency
More like this →
arXiv AI
Jun 17

Curiosity-Critic: Cumulative Prediction Error Improvement as a Tractable Intrinsic Reward for World Model Training

arXiv:2604. 18701v3 Announce Type: replace-cross Abstract: Local prediction-error-based curiosity rewards focus on the current transition without considering the world model's cumulative prediction error across all visited transitions.

By Vin Bhaskara, Haicheng Wang
efficiency
More like this →
arXiv AI
Aug 6

Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability

arXiv:2503. 14833v2 Announce Type: replace-cross Abstract: One of the bottlenecks in robotic intelligence is the instability of neural network models.

By Zihao Liu, Xing Liu, Yuhang Dong, Haitao Chang, Zhengxiong Liu, Panfeng Huang
diffusionroboticsefficiencysafety
More like this →