arXiv Machine Learning

Reinforcement Learning and Rule-Based Peer-to-Peer Pricing in Residential PV-BES Communities

The paper compares rule‑based and reinforcement‑learning (RL) pricing mechanisms for peer‑to‑peer electricity trading in residential photovoltaic communities. Rule‑based benchmarks—bill‑sharing, mid‑market rate, and supply‑demand‑ratio pricing—outperform the best RL policy in a PV‑only setup, while RL policies achieve higher community savings when battery storage is added. Across both configurations, SDR‑shaped pricing outperforms multiplier‑based parameterization, but benefit distribution remains heterogeneous among households.

arXiv AI
Sep 7

Reinforcement Learning for Sequential Solar PV Policy Design under Uncertainty: An Agent-Based Approach

The paper presents a reinforcement learning framework for designing solar PV adoption policies under uncertainty, integrating RL with a stochastic agent‑based model to simulate yearly adoption over a 16‑year horizon. Policymakers can choose annual incentives such as grants, subsidised loans, and feed‑in tariffs, and the study evaluates three RL algorithms—PPO, SAC, and TD3—within a scalarised reward framework that balances adoption gains against costs. Results show clear trade‑off patterns, with TD3 yielding the highest adoption at higher cost, PPO achieving the lowest cost with fewer adopters, and a balanced PPO policy offering a middle ground, all outperforming static baseline policies.

By Iias Faiud, Jonaid Shianifar, Michael Schukat, Karl Mason
arXiv AI
Jun 2

Explainable Data-driven Deep Reinforcement Learning Methods for Optimal Energy Management in Buildings

arXiv:2606. 02049v1 Announce Type: new Abstract: The increasing integration of renewable energy sources into power systems, particularly in buildings equipped with photovoltaic (PV) panels and energy storage systems, introduces significant complexity in energy systems.

By Hallah Shahid Butt, Qiong Huang, G\"okhan Demirel, Kevin F\"orderer, Erfan Tajalli-Ardekani, Simnon Waczowicz, Luigi Spatafora, Veit Hagenmeyer, Benjamin Sch\"afer
arXiv AI
Jul 1

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

arXiv:2606. 31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, voltage fluctuations, line overloads, and the integration of variable renewable energy sources.

By Xavier Rate, Eloann Le Guern, Rapha\"el F\'eraud, Fatma Salem, Melissa Chiknoun, Eymeric Giabicani, Mehdi Feki, Patrick Maill\'e, Guy Camilleri, Anne Blavette, Hamid Benhamed
arXiv Machine Learning
Sep 14

Towards Sustainable Hydrogen Systems: Supply Chain Optimization with Model Predictive Control and Reinforcement Learning

This study evaluates four control strategies—rule-based, model predictive control (MPC), reinforcement learning without forecasts (RL‑NF), and reinforcement learning with forecasts (RL‑F)—for a renewable‑powered hydrogen supply chain. Using a unified, physically realistic simulation that includes electrolyzer constraints, storage dynamics, and grid limits, the authors find that MPC delivers the best economic performance by leveraging short‑term forecasts, while RL‑NF performs robustly without future information. RL‑F does not consistently outperform RL‑NF, indicating that forecast uncertainty and added state complexity can hinder forecast‑augmented learning.

By Mahammad Valiyev
arXiv AI
Jul 15

Constraint-Aware Aggregation for Federated Reinforcement Learning in Microgrid Energy Coordination

arXiv:2607. 12763v1 Announce Type: cross Abstract: Federated Reinforcement Learning (FedRL) enables coordination of distributed energy resources without sharing raw local data, but standard aggregation methods such as FedAvg do not account for system-level constraints, often leading to unsafe global behavior.

By Usman Haider, Karl Mason
arXiv AI
Sep 4

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

The paper introduces DR‑Gym, an open‑source, Gymnasium‑compatible environment that simulates electric utility demand‑response programs at the market level. It uses a regime‑switching wholesale price model calibrated to real extreme events and physics‑based building demand profiles, providing a rich observational space and a configurable multi‑objective reward function for reinforcement learning. Baseline strategies and data snapshots demonstrate the simulator’s realism and learnability.

By Jose E. Aguilar Escamilla, Lingdong Zhou, Xiangqi Zhu, Huazheng Wang
arXiv Machine Learning
Aug 24

BIPPO: Budget-Aware Independent PPO for Energy-Efficient Federated Learning Services

BIPPO (Budget-aware Independent Proximal Policy Optimization) is a multi‑agent reinforcement learning framework designed for energy‑efficient client selection in federated learning (FL) over IoT systems. It addresses infrastructure constraints such as limited resources and device churn, which traditional FL and RL approaches overlook. Evaluated on two image‑classification tasks with non‑IID data, BIPPO improves mean accuracy over non‑RL methods, standard PPO, and IPPO while consuming only a negligible portion of the budget, even as client numbers grow.

By Anna Lackinger, Andrea Morichetta, Pantelis A. Frangoudis, Schahram Dustdar
arXiv Machine Learning
Jun 25

Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources

arXiv:2606. 24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challenged by their inherent uncertainties and modelling complexity.

By Haoyuan Deng, Yihong Zhou, Thomas Morstyn, Yi Wang
arXiv AI
Jul 20

Robustness of Reinforcement Learning-Based Congestion Management in Low-Voltage Grids

arXiv:2607. 16004v1 Announce Type: cross Abstract: Increases in photovoltaic generation, charging of electric vehicles and heat-pump demand challenge operating limits in low-voltage distribution grids.

By Josef Hoppe, Sarra Bouchkati, Farah Nasr, Jonathan Krapp, Alexander Och, Maximilian Wirth, Jan Schiefelbein-Lach, Oliver Pohl, Andreas Ulbig, Michael T. Schaub