arXiv AI

Finite-Sample Probabilistic Safety Certification for AI-Based Grid-Edge Coordination

The paper introduces a finite‑sample probabilistic safety certification framework for black‑box AI decision models used in closed‑loop grid operation. It transforms the AI‑grid evaluation into a binary unsafe outcome under a safety specification and applies exact binomial inference to provide a tight one‑sided upper bound on the unsafe operation probability, using held‑out calibration scenarios. The framework also incorporates physically interpretable sample‑space adversarial attacks to address distribution shifts and is validated through case studies involving 1,000‑agent AI models for grid‑edge flexibility coordination.

Hugging Face Trending Papers
Jun 17

Formal Verification of Learned Multi-Agent Communication Policies via Decision Tree Distillation

Multi-agent reinforcement learning (MARL) enables agents to develop coordination strategies through emergent communication, but neural policies lack the formal safety guarantees required for safety-critical robotic deployment in drone swarms and autonomous vehicle fleets. We present the first end-to-end framework for safety verification of learned multi-agent communication policies through policy abstraction: neural policies are distilled into interpretable decision trees, then formally verified, with empirical validation confirming that verified safety properties transfer to original networks.

arXiv AI
Jun 19

Formal Verification of Learned Multi-Agent Communication Policies via Decision Tree Distillation

arXiv:2606. 19632v1 Announce Type: cross Abstract: Multi-agent reinforcement learning (MARL) enables agents to develop coordination strategies through emergent communication, but neural policies lack the formal safety guarantees required for safety-critical robotic deployment in drone swarms and autonomous vehicle fleets.

By Ahmad Farooq, Kamran Iqbal
arXiv AI
Jul 2

Managed Autonomy at Runtime: Gear-Based Safety and Governance for Single- and Multi-Agent Cyber-Physical Systems

arXiv:2607. 00334v1 Announce Type: new Abstract: Autonomous agents, whether LLM-driven software agents or robotic physical agents, face a common class of failure modes when operating without continuous human oversight: safety violations from unverified actions, behavioral instability from unconstrained loops, and continuity loss from unhandled error states.

By Srini Ramaswamy, Wang Miaosheng
arXiv AI
Aug 25

Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing

The paper presents a robust multi‑agent reinforcement learning framework for small unmanned aircraft systems (sUAS) to maintain separation assurance when GPS data is degraded or spoofed. By modeling state observation corruption as a zero‑sum game, the authors derive a closed‑form adversarial perturbation that eliminates iterative inner optimization and can be evaluated in linear time. Integrating this perturbation into a policy‑gradient MARL algorithm yields a counter‑policy that achieves near‑zero collision rates in high‑density simulations even with up to 35% observation corruption, outperforming non‑adversarial baselines.

By Alex Zongo, Filippos Fotiadis, Ufuk Topcu, Peng Wei