arXiv Machine Learning

Quantifying the Energy Floor: Direct Measurement and Replay Buffer Bias in SAC-Based HVAC Control on sbsim

arXiv:2606. 01665v1 Announce Type: new Abstract: We quantify the energy floor -- the minimum achievable cost given action space constraints -- for Soft Actor-Critic (SAC) HVAC control on the sbsim calibrated building simulator.

arXiv Machine Learning
Jul 20

Comparative Field Deployment of Reinforcement Learning and Model Predictive Control for Residential HVAC

arXiv:2510. 01475v2 Announce Type: replace-cross Abstract: Model Predictive Control (MPC) has demonstrated significant performance improvements over today's control methods for residential Heating, Ventilation, and Air Conditioning (HVAC), but deploying MPC often requires substantial engineering effort.

By Ozan Baris Mulayim, Elias N. Pergantis, Levi D. Reyes Premer, Bingqing Chen, Guannan Qu, Kevin J. Kircher, Mario Berg\'es
arXiv AI
Sep 4

Artificial Intelligence for Energy Optimization in Data Centers

The paper reviews 194 papers on using artificial intelligence to optimize data center energy use, coding 63 of them. It finds that most control studies validate only in simulation, none consider water withdrawal or embodied carbon, and savings estimates overlap across methods, preventing ranking. The authors propose CLEAR‑DC, a framework that links control and workload demand through elasticity, reports net benefits, and records energy, carbon, water, embodied share, and validation venue.

By Mohammed Basharath Ullah, Summaiya Unnisa Begum, Mohammed Nadeem Ullah
arXiv AI
Jun 2

Explainable Data-driven Deep Reinforcement Learning Methods for Optimal Energy Management in Buildings

arXiv:2606. 02049v1 Announce Type: new Abstract: The increasing integration of renewable energy sources into power systems, particularly in buildings equipped with photovoltaic (PV) panels and energy storage systems, introduces significant complexity in energy systems.

By Hallah Shahid Butt, Qiong Huang, G\"okhan Demirel, Kevin F\"orderer, Erfan Tajalli-Ardekani, Simnon Waczowicz, Luigi Spatafora, Veit Hagenmeyer, Benjamin Sch\"afer
arXiv AI
Jul 21

Building2Building: A Large Scale Benchmark for Generalizable Real-World Reinforcement Learning

arXiv:2607. 16534v1 Announce Type: cross Abstract: Reinforcement learning (RL) has achieved strong results in control, yet learned policies remain brittle to changes in dynamics, action spaces, observation spaces, or goals, a critical limitation for real-world deployment.

By Vincent Taboga, Justin Veilleux, Doseok Jang, Anushree Rankawat, Pierre-Luc Bacon