arXiv Machine Learning By Hongpeng Cao, Liqun Zhao, Yuliang Gu, Naira Hovakimyan, Lui Sha, Marco Caccamo

Safe Online Learning via Smooth Safety-Structured Policy Composition

Read the original on arXiv Machine Learning →

arXiv:2606. 31320v1 Announce Type: new Abstract: Safe online reinforcement learning requires policies to respect safety constraints while maintaining smooth optimization dynamics.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.