arXiv Machine Learning By So Nakashima, Tetsuya J. Kobayashi

Accelerating Evolutionary Strategy via Rao-Blackwellizing Realization of Uncertain Input

Read the original on arXiv Machine Learning →

arXiv:2608. 02073v1 Announce Type: cross Abstract: We investigate Optimization under Input Uncertainty (OIU), in which the input to the objective function, rather than the objective function itself, is subject to uncertainty.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

arXiv AI
Jun 6

Retry Policy Gradients in Continuous Action Spaces

arXiv:2606. 05888v1 Announce Type: new Abstract: Retry-based objectives such as pass@K and max@K optimize the best return obtained from multiple sampled trajectories, and recent work has shown that they can promote exploration without explicit exploration bonuses.

By Soichiro Nishimori, Paavo Parmas