arXiv Machine Learning

Sample-Efficient Pareto Front Modeling for Energy-Aware Reinforcement Learning Using Bayesian Optimization

arXiv:2607. 03140v1 Announce Type: new Abstract: Industrial automation increasingly demands control strategies that balance operational performance with strict energy efficiency requirements.

Hugging Face Trending Papers
Jun 29

Toward an Energy-Optimized Operation of Data Centers Located in Wind Farms Using Reinforcement Learning

This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance computing (HPC) data centers. We introduce a reproducible fixed-day simulation framework with synthetic wind and price signals and delayed completion feedback, designed to be extensible toward more complex scenarios.

arXiv Machine Learning
Aug 17

Learning to Run Power Networks: Effective AlphaZero-inspired Topological Control

arXiv:2608. 14114v1 Announce Type: new Abstract: As the integration of volatile renewable energy sources increases the strain on modern power grids, the use of Reinforcement Learning (RL) for autonomous topological reconfiguration has emerged as a promising research field to keep strained grids stable and operational.

By Lukas Zetto, Benjamin Sch\"afer, Qiong Huang
arXiv AI
Jun 18

Pareto Q-Learning with Reward Machines

arXiv:2606. 19134v1 Announce Type: cross Abstract: We present Pareto Q-Learning with Reward Machines (PQLRM), a multi-objective reinforcement learning algorithm for tasks whose reward structure is specified by a set of reward machines (RMs).

By Arnaud Lequen, Cl\'ement Legrand-Lixon, L\'eo Sauli\`eres
arXiv AI
Sep 4

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

The paper introduces DR‑Gym, an open‑source, Gymnasium‑compatible environment that simulates electric utility demand‑response programs at the market level. It uses a regime‑switching wholesale price model calibrated to real extreme events and physics‑based building demand profiles, providing a rich observational space and a configurable multi‑objective reward function for reinforcement learning. Baseline strategies and data snapshots demonstrate the simulator’s realism and learnability.

By Jose E. Aguilar Escamilla, Lingdong Zhou, Xiangqi Zhu, Huazheng Wang