arXiv Machine Learning By Zixuan Liu, Fangzheng Wu, Brian Summa, Zizhan Zheng

Robust General Utility for Reinforcement Learning

Read the original on arXiv Machine Learning →

arXiv:2608. 03562v1 Announce Type: new Abstract: Reinforcement learning (RL) with general utility extends classic RL by optimizing an arbitrary utility functional of the policy-induced occupancy measure, thereby enabling a broader range of applications.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.