arXiv Machine Learning By Pritam Dash, Ethan Chan, Nathan P. Lawrence, Karthik Pattabiraman

ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks

Read the original on arXiv Machine Learning →

arXiv:2506. 22423v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) depend on onboard sensors for perception, navigation, and control.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 21

ASGARD: Action-Space Guard for UAV Resilience via Reinforcement Learning

ASGARD is a two‑phase teacher‑student framework that protects reinforcement‑learning controllers for UAVs from action‑space attacks. In the teacher phase, an encoder fuses the UAV’s physical state with privileged attack information to generate an action‑attack‑aware latent representation, which trains both the control policy and a monitor that corrects actions before they reach the actuators. The student phase then learns to replicate the encoder and monitor using only the UAV’s physical state history, enabling on‑board resilience. Experiments show that ASGARD remains effective against various attack scenarios, including unseen and stealthy attacks, allowing UAV missions to complete successfully.

By Mohsen Salehi, Karthik Pattabiraman
arXiv AI
Jul 3

Lightweight Safe Reinforcement Learning for End-to-End UAV Navigation

arXiv:2607. 01794v1 Announce Type: cross Abstract: With the rapid development of autonomous aerial systems, Unmanned Aerial Vehicles (UAVs) are increasingly deployed in applications such as inspection, environmental monitoring, and rescue, creating growing demand for reliable autonomous navigation.

By Shenghui Zhang, YuXuan Gao, Songwei Zhao, Jifeng Hu, Zijing Zhang, Hechang Chen
arXiv AI
Aug 25

Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing

The paper presents a robust multi‑agent reinforcement learning framework for small unmanned aircraft systems (sUAS) to maintain separation assurance when GPS data is degraded or spoofed. By modeling state observation corruption as a zero‑sum game, the authors derive a closed‑form adversarial perturbation that eliminates iterative inner optimization and can be evaluated in linear time. Integrating this perturbation into a policy‑gradient MARL algorithm yields a counter‑policy that achieves near‑zero collision rates in high‑density simulations even with up to 35% observation corruption, outperforming non‑adversarial baselines.

By Alex Zongo, Filippos Fotiadis, Ufuk Topcu, Peng Wei
Hugging Face Trending Papers
Jul 26

Anticipatory Risk-Guided Reinforcement Learning for Safe Flight Through Dynamic Clutter

Safe quadrotor navigation in cluttered and dynamic environments depends not only on instantaneous geometric perception, but more critically on anticipating collision risks induced by relative motion. Conventional modular pipelines frequently suffer from perception latency, while end-to-end learning methods relying on implicit scalar rewards often struggle to extract reliable spatio-temporal features without physics-grounded supervision.