The paper presents a reinforcement learning framework for designing solar PV adoption policies under uncertainty, integrating RL with a stochastic agent‑based model to simulate yearly adoption over a 16‑year horizon. Policymakers can choose annual incentives such as grants, subsidised loans, and feed‑in tariffs, and the study evaluates three RL algorithms—PPO, SAC, and TD3—within a scalarised reward framework that balances adoption gains against costs. Results show clear trade‑off patterns, with TD3 yielding the highest adoption at higher cost, PPO achieving the lowest cost with fewer adopters, and a balanced PPO policy offering a middle ground, all outperforming static baseline policies.
By Iias Faiud, Jonaid Shianifar, Michael Schukat, Karl Mason
arXiv:2606. 02049v1 Announce Type: new Abstract: The increasing integration of renewable energy sources into power systems, particularly in buildings equipped with photovoltaic (PV) panels and energy storage systems, introduces significant complexity in energy systems.
By Hallah Shahid Butt, Qiong Huang, G\"okhan Demirel, Kevin F\"orderer, Erfan Tajalli-Ardekani, Simnon Waczowicz, Luigi Spatafora, Veit Hagenmeyer, Benjamin Sch\"afer
arXiv:2607. 18272v1 Announce Type: cross Abstract: Prosumers equipped with distributed generation and flexible loads form autonomous cyber-physical energy systems that control local resources and participate in local energy markets with minimal human intervention.
By Lukas Peter Wagner, Raoul Bisson, Felix Gehlhoff
arXiv:2606. 31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, voltage fluctuations, line overloads, and the integration of variable renewable energy sources.
By Xavier Rate, Eloann Le Guern, Rapha\"el F\'eraud, Fatma Salem, Melissa Chiknoun, Eymeric Giabicani, Mehdi Feki, Patrick Maill\'e, Guy Camilleri, Anne Blavette, Hamid Benhamed
arXiv:2607. 06489v1 Announce Type: new Abstract: The dairy industry in Ireland has a large potential for the integration of renewable energy and the reduction of carbon emissions.
By Marcos Eduardo Cruz Victorio, Karl Mason
This study evaluates four control strategies—rule-based, model predictive control (MPC), reinforcement learning without forecasts (RL‑NF), and reinforcement learning with forecasts (RL‑F)—for a renewable‑powered hydrogen supply chain. Using a unified, physically realistic simulation that includes electrolyzer constraints, storage dynamics, and grid limits, the authors find that MPC delivers the best economic performance by leveraging short‑term forecasts, while RL‑NF performs robustly without future information. RL‑F does not consistently outperform RL‑NF, indicating that forecast uncertainty and added state complexity can hinder forecast‑augmented learning.
By Mahammad Valiyev
arXiv:2606. 16051v1 Announce Type: cross Abstract: Residential battery energy storage systems (BESS) are increasingly deployed alongside photovoltaic (PV) generation to reduce household energy costs under volatile time-of-use (TOU) tariffs.
By Dawood Butt, Nandor Verba
arXiv:2607. 12763v1 Announce Type: cross Abstract: Federated Reinforcement Learning (FedRL) enables coordination of distributed energy resources without sharing raw local data, but standard aggregation methods such as FedAvg do not account for system-level constraints, often leading to unsafe global behavior.
By Usman Haider, Karl Mason
The paper introduces DR‑Gym, an open‑source, Gymnasium‑compatible environment that simulates electric utility demand‑response programs at the market level. It uses a regime‑switching wholesale price model calibrated to real extreme events and physics‑based building demand profiles, providing a rich observational space and a configurable multi‑objective reward function for reinforcement learning. Baseline strategies and data snapshots demonstrate the simulator’s realism and learnability.
By Jose E. Aguilar Escamilla, Lingdong Zhou, Xiangqi Zhu, Huazheng Wang
BIPPO (Budget-aware Independent Proximal Policy Optimization) is a multi‑agent reinforcement learning framework designed for energy‑efficient client selection in federated learning (FL) over IoT systems. It addresses infrastructure constraints such as limited resources and device churn, which traditional FL and RL approaches overlook. Evaluated on two image‑classification tasks with non‑IID data, BIPPO improves mean accuracy over non‑RL methods, standard PPO, and IPPO while consuming only a negligible portion of the budget, even as client numbers grow.
By Anna Lackinger, Andrea Morichetta, Pantelis A. Frangoudis, Schahram Dustdar
arXiv:2606. 24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challenged by their inherent uncertainties and modelling complexity.
By Haoyuan Deng, Yihong Zhou, Thomas Morstyn, Yi Wang
arXiv:2607. 16004v1 Announce Type: cross Abstract: Increases in photovoltaic generation, charging of electric vehicles and heat-pump demand challenge operating limits in low-voltage distribution grids.
By Josef Hoppe, Sarra Bouchkati, Farah Nasr, Jonathan Krapp, Alexander Och, Maximilian Wirth, Jan Schiefelbein-Lach, Oliver Pohl, Andreas Ulbig, Michael T. Schaub