arXiv Machine Learning By Shilpa Mukhopadhyay, Sourav Ganguly, Santosh Mohan Rajkumar, Honghao Wei, Debdipta Goswami, Arnob Ghosh

Robust Peak-cost Constrained Reinforcement Learning

Read the original on arXiv Machine Learning →

arXiv:2607. 15457v1 Announce Type: new Abstract: We study robust peak-cost constrained reinforcement learning (RP-CRL), where the objective is to maximize expected reward while controlling the maximum cost encountered along a trajectory.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.