arXiv Machine Learning By Dimitar Chakarov, Lee Cohen, Nathan Srebro

On Incentivized Exploration beyond Bayesianism and Full-Information

Read the original on arXiv Machine Learning →

arXiv:2607. 18300v1 Announce Type: cross Abstract: We extend Incentive Compatible Exploration beyond the Bayesian full-information setting of Kremer et al.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
1d ago

When Do Intrinsic Rewards Lead to Exploration?

The paper investigates when intrinsic rewards effectively drive exploration in reinforcement learning. It introduces a formal criterion that evaluates policies based on the counterfactual information they acquire, comparing how well their histories can replace experience from alternative policies. Using a simple environment, the authors show that common intrinsic reward objectives—count-based, prediction-error, empowerment, and information-gain—can lead to Pareto-suboptimal exploration under this criterion, and they propose conditions and a new objective that better align with optimal exploration.

By Scott W. Viteri (Stanford University), Laura Gomezjurado Gonzalez (Stanford University), Clark Barrett (Stanford University)
arXiv AI
Sep 3

When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection

The paper investigates when information sharing enhances decentralized discovery by separating its effects on pooled estimation and independent rescue actions in finite discovery models. It shows that a registered incremental-sharing protocol improves discovery only when pooled residual error decreases faster than an independent rescue attempt, and that equilibrium selection can determine whether sharing is beneficial. The study uses synthetic, finite models without human or organizational data.

By Yohei Nakajima