arXiv:2606. 02049v1 Announce Type: new Abstract: The increasing integration of renewable energy sources into power systems, particularly in buildings equipped with photovoltaic (PV) panels and energy storage systems, introduces significant complexity in energy systems.
By Hallah Shahid Butt, Qiong Huang, G\"okhan Demirel, Kevin F\"orderer, Erfan Tajalli-Ardekani, Simnon Waczowicz, Luigi Spatafora, Veit Hagenmeyer, Benjamin Sch\"afer
PowerZooJax is a JAX-based benchmark suite designed for reinforcement learning in power system operation. It offers five constrained Markov decision process tasks covering generation, transmission, distribution, distributed energy resources, and data center microgrids. By implementing power flow, economic dispatch, market clearing, and device dynamics as JAX computation graphs, the entire training and evaluation loop runs on the GPU, yielding significant speedups over CPU-based simulations and enabling standardized evaluation of policy returns, safety violations, and out-of-distribution stress conditions.
By Zhanhua Pan, Xiao Liu, Zhilong Cao, Jianhong Wang, Dawei Qiu
arXiv:2607. 12763v1 Announce Type: cross Abstract: Federated Reinforcement Learning (FedRL) enables coordination of distributed energy resources without sharing raw local data, but standard aggregation methods such as FedAvg do not account for system-level constraints, often leading to unsafe global behavior.
By Usman Haider, Karl Mason
This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance computing (HPC) data centers. We introduce a reproducible fixed-day simulation framework with synthetic wind and price signals and delayed completion feedback, designed to be extensible toward more complex scenarios.
arXiv:2608. 15041v1 Announce Type: new Abstract: Coordinating multiple interacting units in complex engineering systems is challenging when system interactions are difficult to model, operational information is heterogeneous, and low-level actions must satisfy strict constraints.
By Changhong He, Jinda Gao, Xinkuan Liu, Le Zhang, Xizi Luo, Yu Mei
This study evaluates four control strategies—rule-based, model predictive control (MPC), reinforcement learning without forecasts (RL‑NF), and reinforcement learning with forecasts (RL‑F)—for a renewable‑powered hydrogen supply chain. Using a unified, physically realistic simulation that includes electrolyzer constraints, storage dynamics, and grid limits, the authors find that MPC delivers the best economic performance by leveraging short‑term forecasts, while RL‑NF performs robustly without future information. RL‑F does not consistently outperform RL‑NF, indicating that forecast uncertainty and added state complexity can hinder forecast‑augmented learning.
By Mahammad Valiyev
arXiv:2505. 05203v3 Announce Type: replace-cross Abstract: With the increasing penetration of renewable energy and inverter-based resources, traditional physics-based power-system operation faces growing challenges in maintaining economic efficiency, security, and robustness.
By Wangkun Xu, Zhongda Chu, Fei Teng
arXiv:2608.28878v1 Announce Type: cross
Abstract: This paper develops a hybrid offline-online multi-agent reinforcement learning framework based on decision transformers. The policy is first pretrain...
By Yiming Zhang, Kun Yang, Cong Shen, Dongning Guo
arXiv:2605. 31044v2 Announce Type: replace Abstract: Reinforcement learning has shown promising results for optimizing the control of industrial energy systems, yet most existing studies remain limited to the application in simulation environments.
By Tobias Lademann, Th\'eo Vincent, Jan Peters, Matthias Weigold
arXiv:2606. 31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, voltage fluctuations, line overloads, and the integration of variable renewable energy sources.
By Xavier Rate, Eloann Le Guern, Rapha\"el F\'eraud, Fatma Salem, Melissa Chiknoun, Eymeric Giabicani, Mehdi Feki, Patrick Maill\'e, Guy Camilleri, Anne Blavette, Hamid Benhamed
arXiv:2607. 27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents.
By Cong Li, Peixi Peng, Yisen Zhao, Xinyu Hu, Shudong Liu, Zhan Su, Zhuojian Li
arXiv:2606. 30316v1 Announce Type: new Abstract: This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance computing (HPC) data centers.
By Jan Stenner, Alexander Kilian, Sebastian Peitz, Hermann de Meer