arXiv AI

LLM-Enhanced Multi-Agent Reinforcement Learning for Unified Electric Vehicles-Charging Station-Grid Optimization in Public Charging Systems

The paper introduces a Large Language Model–enhanced Multi-Agent Reinforcement Learning framework for optimizing electric vehicle charging, station profitability, and grid stability in public charging systems. By using an LLM to select interpretable features from IoT data and dynamically balance conflicting objectives, the approach unifies grid, EV, and station optimization in a single loop. Experiments show the method outperforms existing baselines, improving market efficiency and cutting training time by more than 70%.

arXiv AI
Jul 1

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

arXiv:2606. 31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, voltage fluctuations, line overloads, and the integration of variable renewable energy sources.

By Xavier Rate, Eloann Le Guern, Rapha\"el F\'eraud, Fatma Salem, Melissa Chiknoun, Eymeric Giabicani, Mehdi Feki, Patrick Maill\'e, Guy Camilleri, Anne Blavette, Hamid Benhamed
arXiv Machine Learning
Aug 24

BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services

BIPPO (Budget-aware Independent Proximal Policy Optimization) is a multi‑agent reinforcement learning framework designed for energy‑efficient client selection in federated learning (FL) over IoT systems. It addresses infrastructure constraints such as limited resources and device churn, which traditional FL and RL approaches overlook. Evaluated on two image‑classification tasks with non‑IID data, BIPPO improves mean accuracy over non‑RL methods, standard PPO, and IPPO while consuming only a negligible portion of the budget, even as client numbers grow.

By Anna Lackinger, Andrea Morichetta, Pantelis A. Frangoudis, Schahram Dustdar
arXiv Machine Learning
Aug 27

Simulating Cognitive Smart Freight Corridors with Agent-Based Models and Reinforcement Learning

The paper introduces an agent‑based modeling framework that integrates a physical infrastructure layer, a V2X connectivity layer, and a decision layer using reinforcement learning and multi‑agent reinforcement learning to simulate smart freight corridors. Three scenarios—Baseline, Assisted, and Cognitive—are evaluated on throughput, congestion, energy, emissions, and robustness, with the Cognitive scenario outperforming the baseline in throughput and congestion, and the Assisted scenario achieving energy savings via platooning. Sensitivity analysis shows that the smart corridor’s throughput advantage grows under high demand and that MARL coordination better utilizes fixed charging capacity than rule‑based methods.

By Madelaine Martinez-Ferguson, Chun Wang, Mustafa Can Camur, Xueping Li
arXiv AI
Sep 4

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

The paper introduces DR‑Gym, an open‑source, Gymnasium‑compatible environment that simulates electric utility demand‑response programs at the market level. It uses a regime‑switching wholesale price model calibrated to real extreme events and physics‑based building demand profiles, providing a rich observational space and a configurable multi‑objective reward function for reinforcement learning. Baseline strategies and data snapshots demonstrate the simulator’s realism and learnability.

By Jose E. Aguilar Escamilla, Lingdong Zhou, Xiangqi Zhu, Huazheng Wang
arXiv AI
Jun 17

Enhanced Evolutionary Multi-Objective Deep Reinforcement Learning for Reliable and Efficient Wireless Rechargeable Sensor Networks

arXiv:2510. 21127v2 Announce Type: replace-cross Abstract: Despite rapid advancements in sensor networks, conventional battery-powered sensor networks suffer from limited operational lifespans and frequent maintenance requirements that severely constrain their deployment in remote and inaccessible environments.

By Bowei Tong, Hui Kang, Jiahui Li, Geng Sun, Jiacheng Wang, Yaoqi Yang, Bo Xu, Dusit Niyato
arXiv Machine Learning
3d ago

OpenHail: An Event-Driven Gymnasium Environment for Electric Ride-Hailing Fleet Control

OpenHail is an open-source Gymnasium environment designed for controlling electric ride‑hailing fleets. It offers a fixed‑size observation–action interface that handles request assignment, repositioning, and charging, while its event‑driven simulator models pickup deadlines, vehicle job queues, battery dynamics, and finite‑capacity charging facilities with FIFO queues. The environment supports various decision‑epoch mechanisms—event‑driven, periodic, hybrid, and policy‑requested—allowing flexible policy interactions within a unified operational model, and includes tools for evaluation, metrics, and baseline policies.

By Tommaso Schettini, Nicholas D. Kullman, Jorge E. Mendoza
arXiv Machine Learning
Jun 25

Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources

arXiv:2606. 24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challenged by their inherent uncertainties and modelling complexity.

By Haoyuan Deng, Yihong Zhou, Thomas Morstyn, Yi Wang
arXiv Machine Learning
Jun 10

Toward Proactive RF Charging Scheduling: Generative AI for Decision Support

arXiv:2606. 10600v1 Announce Type: cross Abstract: Radio frequency wireless power transfer (RF-WPT) is an enabling technology for supporting uninterrupted communications in future Internet of Things systems by reducing the need for battery replacement and mitigating battery-waste-related issues.

By Amirhossein Azarbahram, Osmel M. Rosabal, David Ernesto Ruiz-Guirola, Melike Erol-Kantarci, Kaibin Huang, Onel L. A. L\'opez
arXiv AI
Aug 28

SynthCharge: An Electric Vehicle Routing Instance Generator with Feasibility Screening to Enable Learning-Based Optimization and Benchmarking

SynthCharge is a parametric generator that creates diverse, feasibility‑screened instances of the electric vehicle routing problem with time windows (EVRPTW). It produces instances ranging from 5 to 100 customers (up to 500 in theory) with adaptive energy capacity scaling and range‑aware charging station placement, filtering out unsolvable cases via a fast feasibility screening process. This dynamic benchmarking infrastructure enables systematic evaluation of learning‑based routing and data‑driven approaches.

By Mertcan Daysalilar, Fuat Uyguroglu, Gabriel Nicolosi, Adam Meyers
arXiv AI
Jun 2

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control

arXiv:2604. 17456v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown strong capabilities in long-horizon reasoning, tool use, and decision-making in digital environments, yet extending them to physically grounded systems remains challenging.

By Siqi Lai, Pan Zhang, Yuping Zhou, Jindong Han, Yansong Ning, Hao Liu