arXiv AI

Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles

arXiv:2607. 08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from normal operation before they propagate into failures.

Hugging Face Trending Papers
Jul 14

OOD-RL-Bench: A Benchmark Framework for Out-of-Distribution Detection in Reinforcement Learning

Reliable reinforcement learning (RL) agents must maintain operational integrity amidst sensor malfunctions, dynamic disturbances, and slow environmental shifts. The detection of out-of-distribution conditions is pivotal to determining when an agent's observations, transitions, or trajectory dynamics deviate from the assumptions underpinning its policy training.

arXiv AI
4d ago

MAADBench: The Refreshable Paradigm for Anomaly Detection in Multi-Agent Systems

MAADBench is a refreshable benchmark for anomaly detection in multi‑agent systems powered by large language models. It addresses the challenge of keeping benchmarks current by sampling and coupling generative tasks, generating trace data under configurable LLM backbones, and automatically providing deterministic step‑level labels. The authors evaluated 25 anomaly‑detection methods on 5,200 labeled traces, finding that existing approaches depend heavily on supervision, struggle with subtle MAS‑specific anomalies, and lack robustness across different LLM backbones.

By Lei Ma, Dennis Hofmann, Haowen Xu, Joshua DeOliveira, Peter VanNostrand, Lei Cao, Elke Rundensteiner
arXiv Machine Learning
Sep 7

Online Change-point Detection for Cooperative Multi-Agent Reinforcement Learning

The paper introduces Patterns of Past Rewards (PPR), a lightweight, algorithm‑agnostic online change‑point detector for cooperative multi‑agent reinforcement learning. PPR smooths agents’ return streams, highlights recent changes, and applies a statistical drift detector to flag significant shifts. Experiments in a custom Speaker‑Listener environment show that PPR balances detection speed and alarm stability, outperforming both a smoothed‑return baseline and a raw‑return detector.

By Fatemeh Saberi Khomami, Julita Vassileva
arXiv AI
Aug 13

Machine Learning-Based Cyber Defense for Cloud Infrastructure: An Adaptive Deep Q-Network Architecture for Intelligent Intrusion Detection and Automated Threat Mitigation

arXiv:2608. 12190v1 Announce Type: cross Abstract: With the increasing complexity of cyber assaults in cloud environments, adaptable security solutions are needed that can support real-time detection and autonomous response.

By Md Yassir Mottalib, Md Yousuf, Eklachur Rahman Bhuiyan, S M Ahsan Habib, Sonjoy Kumar Dey, Md. Salahuddin Gazi, Molay Kumar Roy, Asaduzzaman Anik