← Back to all news
OpenAI Blog March 3, 2018

Some considerations on learning to explore via meta-reinforcement learning

Read the original on OpenAI Blog →

The Flow has not summarised this story yet — read it at OpenAI Blog.

  • reinforcement-learning

Related stories

OpenAI Blog
Aug 13, 2018

Large-scale study of curiosity-driven learning

More like this →
arXiv AI
Jun 19

Can In-Context Learning Support Intrinsic Curiosity?

arXiv:2606. 19476v1 Announce Type: cross Abstract: Effective machine learning depends not only on how we model data, but also on what data we choose to collect.

By Eric Elmoznino, Sangnie Bhardwaj, Johannes von Oswald, Rajai Nasser, Blaise Ag\"uera y Arcas, Jo\~ao Sacramento, Rif A. Saurous, Guillaume Lajoie
llmsagentsreinforcement-learningsafety
More like this →
arXiv AI
Jul 21

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning

arXiv:2607. 16421v1 Announce Type: new Abstract: It has long been recognized that humans have the ability to switch between fast, reactive decision-making and slower, deliberative planning.

By Adam Labiosa, Josiah P. Hanna
agentsreinforcement-learning
More like this →
arXiv AI
2d ago

Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents

arXiv:2608. 14339v1 Announce Type: new Abstract: We study proactive exploration in LLM agents, i.

By Zhizhao Guan, Chen Huang, Ziming Liu, Hongru Liang, Wenqiang Lei, See-Kiong Ng, Tat-Seng Chua, Anthony G Cohn
llmsagentsreinforcement-learningsafety
More like this →
OpenAI Blog
Nov 5, 2018

Plan online, learn offline: Efficient learning and exploration via model-based control

More like this →
Hugging Face Trending Papers
Jun 17

Can In-Context Learning Support Intrinsic Curiosity?

Effective machine learning depends not only on how we model data, but also on what data we choose to collect. While large sequence models have revolutionized data modeling, the problem of automated data selection, or "intrinsic curiosity", remains a significant challenge.

llmsagentsreinforcement-learningsafety
More like this →