arXiv Machine Learning By Ziyuan Wang (Steven), Yifan Sui (Steven), Wei Wei (Steven), Wenjie Xin (Steven), Zekai Zhang (Steven), Xiangwang Hou (Steven), Xiao-Ping (Steven), Zhang

AERIS: Offline Policy Improvement for Multi-UAV Integrated Sensing and Communication

Read the original on arXiv Machine Learning →

AERIS is an offline policy improvement framework for multi-UAV integrated sensing and communication (ISAC) that learns from fixed flight logs using centralized training and decentralized execution. It introduces STAR-CRDT, an offline multi-agent RL algorithm that rectifies local actions and distills trusted improvements into decentralized actors, providing an offline-support policy improvement guarantee. Experiments demonstrate that STAR-CRDT boosts the main ISAC objective return by 29.3% and improves communication sum rate, sensing pass rate, and sensing margin while reducing collision-risk events by 54.2%.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Aug 7

Communication-Aware Multi-Agent Reinforcement Learning for Decentralized Cooperative UAV Deployment

arXiv:2603. 16141v2 Announce Type: replace-cross Abstract: Autonomous Unmanned Aerial Vehicle (UAV) swarms are increasingly used as rapidly deployable aerial relays and sensing platforms, yet practical deployments must operate under partial observability and intermittent peer-to-peer connectivity.

By Enguang Fan, Yifan Chen, Zihan Shan, Klara Nahrstedt, Matthew Caesar, Jae Kim