arXiv AI

LLM-Empowered Agentic MAC Protocols: A Dynamic Stackelberg Game Approach

arXiv:2510. 10895v2 Announce Type: replace Abstract: Medium Access Control (MAC) protocols, essential for wireless networks, are typically manually configured.

arXiv AI
Sep 18

Agentic AI Networking for Heterogeneous Unmanned Aerial Systems in Low-Altitude Wireless Networks

The paper introduces a hierarchical hybrid architecture combining large language models (LLMs) and multi-agent reinforcement learning (MARL) to manage heterogeneous unmanned aerial systems in low‑altitude wireless networks (LAWNs). An outer LLM‑driven loop interprets service requirements and operator intent to reconfigure objectives and resource priorities, while an inner MARL loop executes decentralized policies under the updated game. A logistics‑monitoring case study demonstrates the framework’s ability to coordinate diverse services and adapt to changing conditions without retraining the MARL policies.

By Nguyen Duc Minh Quang, Chang Liu, Shuangyang Li, Derrick Wing Kwan Ng
arXiv AI
Sep 4

From Prior-Guided Heuristics to Deployable Agents: Accelerating Demonstration-Driven Reinforcement Learning for Deadline-Constrained Network Control

The paper proposes a deployment‑focused framework for deadline‑constrained network control, introducing the Effective Congestion (EC) metric family and Uniform Path Grouping (UPG) heuristic to better capture traffic urgency and balance load. It integrates these with a Multi‑Agent Deep Reinforcement Learning architecture (MADRL EC (p*)) that combines a distributed scheduler and a centralized RL router. A unified training objective merges live‑reward, pre‑collected‑reward, and policy‑imitation terms, leading to the Model‑Guided Annealed Reinforcement Learning (MGA‑RL) protocol built on DDPG, which generalizes offline‑to‑online learning for demonstration‑driven training.

By Vincenzo Norman Vitale, Mohammad Solki, Antonia Maria Tulino, Andreas F. Molisch, Jaime Llorca
arXiv AI
Jun 4

Generalizable Multi-Task Learning for Wireless Networks Using Prompt Decision Transformers

arXiv:2606. 04328v1 Announce Type: cross Abstract: Future wireless networks demand rapid adaptation to highly heterogeneous environments and dynamic task configurations, necessitating a shift from conventional rule-based and optimization-driven radio resource management (RRM) toward artificial intelligence (AI)-driven RRM.

By Fatih Temiz, Shavbo Salehi, Melike Erol-Kantarci
arXiv Machine Learning
Jun 19

Utility-Aware DRL-Based TXOP Adaptation for NR-U and Wi-Fi Coexistence Networks

arXiv:2605. 00457v4 Announce Type: replace-cross Abstract: The coexistence of NR-U and Wi-Fi in the unlicensed spectrum introduces a challenging resource management problem, where heterogeneous channel access mechanisms can lead to unbalanced spectrum utilization and severe Wi-Fi performance degradation.

By Po-Heng Chou, Yi-Fang Yu, Shou-Yu Chen, Chiapin Wang
arXiv AI
Jun 2

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control

arXiv:2604. 17456v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown strong capabilities in long-horizon reasoning, tool use, and decision-making in digital environments, yet extending them to physically grounded systems remains challenging.

By Siqi Lai, Pan Zhang, Yuping Zhou, Jindong Han, Yansong Ning, Hao Liu
arXiv AI
Aug 24

FL-MAESTRO: Multi-Agent LLM Orchestration for Resource-Constrained Federated Learning

FL-MAESTRO is a multi‑agent orchestrator that uses three specialized large language model agents to jointly decide the communication topology, per‑client resource allocation, and aggregation rule in each federated learning round. A coordinator merges the agents’ analyses, and a non‑LLM feasibility check validates the decision before execution. By filtering out clients whose updates would never be aggregated, the system eliminates the main source of wasted round energy in volatile edge networks and works across heterogeneous device classes without per‑class energy models, achieving comparable accuracy to the best energy‑aware baseline while reducing wasted energy from over a third to near zero on a non‑IID CIFAR‑10 benchmark.

By Jiajun Wu, Zirui Wang, Jiayu Zhou, Qiang Ye, Steve Drew
arXiv AI
Jul 7

Multi-Agent Reinforcement Learning for V2X Resource Allocation: Disentangling MARL Challenges Through Benchmarking

arXiv:2603. 06607v2 Announce Type: replace-cross Abstract: Radio resource allocation (RRA) is a critical function in cellular vehicle-to-everything (C-V2X) networks, where vehicles must share limited wireless resources to support safety-critical communications.

By Siyuan Wang, Lei Lei, Pranav Maheshwari, Sam Bellefeuille, Kan Zheng
arXiv Machine Learning
Aug 5

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

arXiv:2608. 03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource control across independently deployable cell-level controllers in open and disaggregated 6G RANs.

By Amin Farajzadeh, Melike Erol-Kantarci