arXiv Machine Learning

Learn for Variation: Efficient AAV Trajectory Learning through a Differentiable Wireless World Model

arXiv:2603. 18853v3 Announce Type: replace-cross Abstract: Autonomous aerial vehicles (AAVs) enable data collection for sixth-generation Internet-of-Things networks, but their trajectories couple nonlinear wireless rates with long-horizon service progress.

arXiv AI
Aug 26

Place, Slice and Schedule: Hierarchical O-RAN Control of a Tethered mmWave UAV-gNB

The paper proposes a hierarchical Open Radio Access Network (O‑RAN) control framework for tethered millimeter‑wave UAV‑mounted 5G base stations (gNBs). A Non‑Real‑Time RIC application jointly manages UAV placement and slice budgets, while a Near‑Real‑Time RIC application allocates per‑user resources using a permutation‑equivariant DeepSets Soft Actor‑Critic scheduler. This two‑level controller improves eMBB service‑level agreement satisfaction by up to 17 % and URLLC on‑time delivery by up to 42 % compared with conventional schedulers.

By Alireza Mohammadhosseini, Fatemeh Afghah
arXiv AI
Jun 19

Oranits: Mission Assignment and Task Offloading in Open RAN-based ITS using Metaheuristic and Deep Reinforcement Learning

arXiv:2507. 19712v3 Announce Type: replace-cross Abstract: In this paper, we explore mission assignment and task offloading in an Open Radio Access Network (Open RAN)-based intelligent transportation system (ITS), where autonomous vehicles leverage mobile edge computing for efficient processing.

By Ngoc Hung Nguyen, Nguyen Van Thieu, Quang-Trung Luu, Anh Tuan Nguyen, Senura Wanasekara, Nguyen Cong Luong, Fatemeh Kavehmadavani, Van-Dinh Nguyen
arXiv AI
Jul 21

Lyapunov Stability-Aware Stackelberg Game for Low-Altitude Economy: A Control-Oriented Pruning-Based DRL Approach

arXiv:2602. 01131v2 Announce Type: replace Abstract: With the rapid expansion of the low-altitude economy, Unmanned Aerial Vehicles (UAVs) serve as pivotal aerial base stations supporting diverse services from users, ranging from latency-sensitive critical missions to bandwidth-intensive data streaming.

By Yue Zhong, Jiawen Kang, Yongju Tong, Hong-Ning Dai, Dong In Kim, Abbas Jamalipour, Shengli Xie
arXiv AI
Jun 3

AUGUSTE: Online-Learning dApp for Predictive URLLC Scheduling

arXiv:2606. 03664v1 Announce Type: cross Abstract: Ultra Reliable and Low Latency Communications (URLLC) was one of the main motivations behind 5G, with 3GPP advertising 1-10 ms latency targets for applications such as industrial automation, Vehicle-To-Everything (V2X), tactical edge networking, and unmanned-system control.

By Maxime Elkael, Michele Polese, Yunseong Lee, Koichiro Furueda, Tommaso Melodia
arXiv AI
5d ago

Agentic AI Networking for Heterogeneous Unmanned Aerial Systems in Low-Altitude Wireless Networks

The paper introduces a hierarchical hybrid architecture combining large language models (LLMs) and multi-agent reinforcement learning (MARL) to manage heterogeneous unmanned aerial systems in low‑altitude wireless networks (LAWNs). An outer LLM‑driven loop interprets service requirements and operator intent to reconfigure objectives and resource priorities, while an inner MARL loop executes decentralized policies under the updated game. A logistics‑monitoring case study demonstrates the framework’s ability to coordinate diverse services and adapt to changing conditions without retraining the MARL policies.

By Nguyen Duc Minh Quang, Chang Liu, Shuangyang Li, Derrick Wing Kwan Ng
arXiv Machine Learning
6d ago

FedPGT: Progressive Gradient Transmission for Vehicular Federated Learning over Time-Varying Channels

FedPGT introduces a progressive gradient transmission scheme for vehicular federated learning over time‑varying channels, where vehicles send high‑magnitude gradient entries according to instantaneous channel conditions. The authors derive a convergence bound showing diminishing returns governed by a power‑law decay, and formulate a stochastic optimization problem that is solved via a Lyapunov drift‑plus‑penalty approach with per‑slot surrogate variables. A low‑complexity resource allocation algorithm is proposed, and experiments on CIFAR‑10 and Argoverse demonstrate a 3.65% accuracy gain and a 12.66% reduction in displacement error compared to state‑of‑the‑art baselines.

By Jintao Yan, Tan Chen, Yuxuan Sun, Sheng Zhou, Zhisheng Niu
arXiv Machine Learning
Sep 2

Accelerating Reinforcement Learning via MPC Solver-Gradient Guidance for Weights-varying MPC

The paper introduces Solver-Gradient Guided Reinforcement Learning (SG‑RL), a method that augments standard RL with bounded gradients from a differentiable MPC solver to adapt cost‑function weights online. SG‑RL integrates solver‑gradient guidance into PPO through actor‑update scaling, policy loss, advantage estimation, and value‑function learning, achieving comparable or superior closed‑loop performance while requiring up to 70.6% fewer samples. Experiments on two autonomous racing platforms with intentional model mismatch demonstrate that SG‑RL outperforms both RL and gradient‑based policy learning baselines and generalizes zero‑shot to unseen environments.

By Baha Zarrouki, Arslan Thobani, Jasper Hoffmann, Mattia Piccinini, Rudolf Reiter, Felix Jahncke, S\'ebastien Gros, Davide Scaramuzza, Johannes Betz