arXiv Machine Learning

Multi-Agent Off-Policy Deep Reinforcement Learning for Smart Campus Coverage

The paper studies optimal placement of millimeter-wave base stations in a realistic, non-convex campus layout using deep reinforcement learning. It compares four DRL methods—single-agent DQN, multi-agent partitioned DQN, single-agent DDPG, and multi-agent partitioned DDPG—and finds that the multi-agent DDPG approach achieves full coverage, a Jain's fairness index of 0.94, and superior performance in dense scenarios with 400 users. The multi-agent DDPG also converges more efficiently than single-agent methods.

arXiv Machine Learning
Jun 19

Utility-Aware DRL-Based TXOP Adaptation for NR-U and Wi-Fi Coexistence Networks

arXiv:2605. 00457v4 Announce Type: replace-cross Abstract: The coexistence of NR-U and Wi-Fi in the unlicensed spectrum introduces a challenging resource management problem, where heterogeneous channel access mechanisms can lead to unbalanced spectrum utilization and severe Wi-Fi performance degradation.

By Po-Heng Chou, Yi-Fang Yu, Shou-Yu Chen, Chiapin Wang
arXiv AI
Jun 4

Generalizable Multi-Task Learning for Wireless Networks Using Prompt Decision Transformers

arXiv:2606. 04328v1 Announce Type: cross Abstract: Future wireless networks demand rapid adaptation to highly heterogeneous environments and dynamic task configurations, necessitating a shift from conventional rule-based and optimization-driven radio resource management (RRM) toward artificial intelligence (AI)-driven RRM.

By Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci
arXiv AI
Sep 10

Learning to Focus: CSI-Free Hierarchical MARL for Reconfigurable Reflectors

The paper proposes a CSI‑free hierarchical multi‑agent reinforcement learning framework for controlling reconfigurable reflective surfaces in millimeter‑wave networks. By replacing per‑element channel estimation with user localization data, the system uses a two‑tier neural architecture: a high‑level controller for discrete user‑to‑reflector assignments and low‑level controllers that optimize continuous focal points via MAPPO under a CTDE scheme. Deterministic ray‑tracing tests show RSSI gains of up to 7.79 dB over centralized PPO baselines and robust performance with sub‑meter localization errors for multiple users and reflector arrays.

By Hieu Le, Mostafa Ibrahim, Oguz Bedir, Jian Tao, Sabit Ekin
arXiv AI
Jun 2

Digital Twin-Assisted Adaptive Multi-Agent DRL for Intelligent Spectrum and Resource Management in Open-RAN UAV-Enabled 6G Networks

arXiv:2606. 01324v1 Announce Type: cross Abstract: The evolution toward 6G wireless networks envisions a seamlessly intelligent, Open-RAN-enabled architecture where unmanned aerial vehicles (UAVs) play a pivotal role in extending coverage, enhancing resilience, and ensuring reliable connectivity for ground users deployment.

By Marwan Dhuheir, Thang X. Vu, Symeon Chatzinotas
arXiv AI
Jun 19

Oranits: Mission Assignment and Task Offloading in Open RAN-based ITS using Metaheuristic and Deep Reinforcement Learning

arXiv:2507. 19712v3 Announce Type: replace-cross Abstract: In this paper, we explore mission assignment and task offloading in an Open Radio Access Network (Open RAN)-based intelligent transportation system (ITS), where autonomous vehicles leverage mobile edge computing for efficient processing.

By Ngoc Hung Nguyen, Nguyen Van Thieu, Quang-Trung Luu, Anh Tuan Nguyen, Senura Wanasekara, Nguyen Cong Luong, Fatemeh Kavehmadavani, Van-Dinh Nguyen
arXiv Machine Learning
Aug 4

Heterogeneous Multi-Agent Reinforcement Learning for Radio Resource Management under Coupled Finite-Horizon Constraints

arXiv:2608. 01745v1 Announce Type: new Abstract: Maximizing throughput under proportional fairness in dense wireless networks requires jointly managing user association, scheduling, base station (BS) activation, and handover control under hard finite-horizon energy and handover budgets, which induces a fundamental tension between BS-side energy management and user-side handover regulation.

By Yeonseo Jeong, Wonhyeok Ko, Sungweon Hong, Songnam Hong
arXiv AI
Jul 7

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

arXiv:2605. 12569v2 Announce Type: replace-cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environments where source localization is highly challenging.

By M. Shamail J. Khan, Nisha L. Raichur, Lucas Heublein, Christian Wielenberg, Alexander Mattick, Tobias Feigl, Christopher Mutschler, Felix Ott