← Back to all news
OpenAI Blog November 15, 2016

#Exploration: A study of count-based exploration for deep reinforcement learning

Read the original on OpenAI Blog →

The Flow has not summarised this story yet — read it at OpenAI Blog.

  • reinforcement-learning

Related stories

arXiv Machine Learning
Jul 21

Information-Based Exploration via Random Features for Reinforcement Learning

arXiv:2607. 17981v1 Announce Type: new Abstract: Representation learning has enabled classical exploration strategies to be extended to deep Reinforcement Learning (RL), but often makes algorithms more complex and theoretical guarantees harder to establish.

By Waris Radji, Odalric-Ambrym Maillard
reinforcement-learning
More like this →
arXiv AI
6d ago

DORA Explorer: Improving the Exploration Ability of LLMs Without Training

arXiv:2604. 17244v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents for sequential decision-making struggle to produce diverse outputs.

By Priya Gurjar, Md Farhan Ishmam, Kenneth Marino
llmsagentsreinforcement-learning
More like this →
arXiv AI
Aug 3

Explore Beyond the Boundary Using Entropic Information

arXiv:2607. 29419v1 Announce Type: cross Abstract: In reinforcement learning, exploration with sparse and delayed rewards presents a significant challenge due to the limited feedback available for guiding the learning process.

By Bumgeun Park, Donghwan Lee
agentsreinforcement-learning
More like this →
OpenAI Blog
Nov 21, 2019

Benchmarking safe exploration in deep reinforcement learning

reinforcement-learningbenchmarks
More like this →
OpenAI Blog
Jul 27, 2017

Better exploration with parameter noise

We’ve found that adding adaptive noise to the parameters of reinforcement learning algorithms frequently boosts performance. This exploration method is simple to implement and very rarely decreases performance, so it’s worth trying on any problem.

reinforcement-learning
More like this →
arXiv AI
2d ago

Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in LLM Agents

arXiv:2608. 14339v1 Announce Type: new Abstract: We study proactive exploration in LLM agents, i.

By Zhizhao Guan, Chen Huang, Ziming Liu, Hongru Liang, Wenqiang Lei, See-Kiong Ng, Tat-Seng Chua, Anthony G Cohn
llmsagentsreinforcement-learningsafety
More like this →