arXiv AI

A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integration for Linear Systems

arXiv:2604. 21030v2 Announce Type: replace-cross Abstract: The integration of Model Predictive Control (MPC) and Reinforcement Learning (RL) has emerged as a promising paradigm for constrained decision-making and adaptive control.

arXiv Machine Learning
Sep 2

Accelerating Reinforcement Learning via MPC Solver-Gradient Guidance for Weights-varying MPC

The paper introduces Solver-Gradient Guided Reinforcement Learning (SG‑RL), a method that augments standard RL with bounded gradients from a differentiable MPC solver to adapt cost‑function weights online. SG‑RL integrates solver‑gradient guidance into PPO through actor‑update scaling, policy loss, advantage estimation, and value‑function learning, achieving comparable or superior closed‑loop performance while requiring up to 70.6% fewer samples. Experiments on two autonomous racing platforms with intentional model mismatch demonstrate that SG‑RL outperforms both RL and gradient‑based policy learning baselines and generalizes zero‑shot to unseen environments.

By Baha Zarrouki, Arslan Thobani, Jasper Hoffmann, Mattia Piccinini, Rudolf Reiter, Felix Jahncke, S\'ebastien Gros, Davide Scaramuzza, Johannes Betz
Hugging Face Trending Papers
Jun 23

Solving Markov Decision Processes with Future Information via MPC

Model Predictive Control (MPC) is widely used in industrial and robotic systems for enforcing constraints and embedding domain knowledge through finite-horizon optimization-based planning. However, despite these strengths, an MPC scheme typically does not yield optimal policies for sequential decision-making problems formulated as Markov Decision Processes (MDPs).

arXiv Machine Learning
Sep 14

Towards Sustainable Hydrogen Systems: Supply Chain Optimization with Model Predictive Control and Reinforcement Learning

This study evaluates four control strategies—rule-based, model predictive control (MPC), reinforcement learning without forecasts (RL‑NF), and reinforcement learning with forecasts (RL‑F)—for a renewable‑powered hydrogen supply chain. Using a unified, physically realistic simulation that includes electrolyzer constraints, storage dynamics, and grid limits, the authors find that MPC delivers the best economic performance by leveraging short‑term forecasts, while RL‑NF performs robustly without future information. RL‑F does not consistently outperform RL‑NF, indicating that forecast uncertainty and added state complexity can hinder forecast‑augmented learning.

By Mahammad Valiyev
arXiv Machine Learning
Sep 17

Composite-Gradient Learning for Shared Control Authority Between Deep Reinforcement Learning and Model Predictive Control

The paper introduces Composite‑Gradient Learning (CGL), a method that explicitly incorporates a model predictive controller (MPC) into the training of a deep reinforcement learning (DRL) agent by treating their control inputs as a joint action. CGL updates the DRL policy while accounting for the interaction with the MPC, unlike prior approaches that view MPC merely as part of the environment. Experiments on two freeway traffic networks show that CGL performs better than alternative methods when the interaction between DRL and MPC is strong, though overall gains are modest.

By Giray \"On\"ur, Azita Dabiri, Bart De Schutter
Hugging Face Trending Papers
Jul 14

Learning-enabled Acceleration of Scenario-based Model Predictive Control

Scenario-based model predictive control (SBMPC) is a variant of model predictive control (MPC) that explicitly accounts for uncertainty by optimizing control actions over multiple predicted scenarios. However, its computational complexity increases rapidly with the number of scenarios and prediction horizon, limiting is applicability to real-time planning and control.