Hugging Face Trending Papers

Predictive safety filter enhanced curriculum learning control for efficient vehicle dynamics controller

Recent advances in learning-based control have enabled impressive achievements in solving complex control problems in various domains. However, since learning-based control may not be able to realize safety-guaranties, it is of great importance to enhance safety and robustness while maintaining good performances.

arXiv AI
Sep 24

Turning Safety into Competence: Minimally Exploitable Robot Policies via Safety-Filtered Reinforcement Learning

The paper introduces Safety to Competence (S2C), a two‑stage reinforcement learning framework that first learns a safety filter and then trains a competitive task policy while embedding the filter. By separating safety synthesis from task learning, S2C reduces training complexity and prevents the policy from being exploited by adversarial attacks. Experiments on simulated touchdown games show that S2C achieves higher win rates, better Elo ratings, and lower exploitability than eight safe‑RL baselines, and hardware tests confirm its competence against a human opponent.

By Ruihan Wu, Rui Yang, Donggeon David Oh, Duy Nguyen, Haimin Hu
arXiv Machine Learning
Jul 23

Safety-Regulated Transfer Reinforcement Learning with Adaptive Teacher Guidance

arXiv:2606. 26527v2 Announce Type: replace Abstract: We propose Safety-Regulated Adaptive Transfer Reinforcement Learning (SRATRL), a teacher--student framework that combines safety-triggered intervention, safety-adaptive value shaping, and policy-compatibility-based optimization for efficient target-domain adaptation.

By Wenjie Huang, Yang Li, Jingjia Teng, Mingwei Jin, Kai Song, Zeyu Yang, Qisong Yang, Yougang Bian
arXiv AI
Sep 17

CALOS: Control-Affine Lyapunov On-manifold Safety Layer for Safe Deep Reinforcement Learning for Quadrotors

CALOS is a runtime safety layer for quadrotor control that enforces attitude constraints without altering the underlying deep reinforcement learning algorithm. It formulates tilt-angle inequalities and a Lyapunov descent condition into a single quadratic program, solved exactly via active-set enumeration over a three-dimensional torque space. In NVIDIA Isaac Lab trajectory-tracking tasks, CALOS reduces lateral tracking error by 55‑60% compared to an unconstrained Proximal Policy Optimization baseline and eliminates attitude-constraint violations during training.

By Fabrizio Cesareo, Sebastiano Mengozzi, Nicola Mimmo, Andrea Acquaviva