paper-with-me

홈 › Papers

MPCritic: A plug-and-play MPC architecture for reinforcement learning

2025-04-01 · Nathan P. Lawrence, Thomas Banker, Ali Mesbah

The reinforcement learning (RL) and model predictive control (MPC) communities have developed vast ecosystems of theoretical approaches and computational tools for solving optimal control problems. Given their conceptual similarities but differing strengths, there has been increasing interest in synergizing RL and MPC. However, existing approaches tend to be limited for various reasons, including computational cost of MPC in an RL algorithm and software hurdles towards seamless integration of MPC and RL tools. These challenges often result in the use of "simple" MPC schemes or RL algorithms, neglecting the state-of-the-art in both areas. This paper presents MPCritic, a machine learning-friendly architecture that interfaces seamlessly with MPC tools. MPCritic utilizes the loss landscape defined by a parameterized MPC problem, focusing on "soft" optimization over batched training steps; thereby updating the MPC parameters while avoiding costly minimization and parametric sensitivities. Since the MPC structure is preserved during training, an MPC agent can be readily used for online deployment, where robust constraint satisfaction is paramount. We demonstrate the versatility of MPCritic, in terms of MPC architectures and RL algorithms that it can accommodate, on classic control benchmarks.

📄 PDF Abstract BibTeX arXiv:2504.01086

Code (1)

tbanker/MPCritic 공식 구현 pytorch

Tasks

Model Predictive ControlReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Soft MPCritic: Amortized Model Predictive Value Iteration

2026-04-01 · Thomas Banker, Nathan P. Lawrence, Ali Mesbah arxiv

Reinforcement learning (RL) and model predictive control (MPC) offer complementary strengths, yet combining them at scale remains computationally challenging. We propose soft MPCritic, an RL-MPC framework that learns in …

Reinforcement Learning

RLFactory: A Plug-and-Play Reinforcement Learning Post-Training Framework for LLM Multi-Turn Tool-Use

2025-08-31 · Jiajun Chai, Guojun Yin, Zekun Xu, Chuhuai Yue 외 arxiv

Large language models excel at basic reasoning but struggle with tasks that require interaction with external tools. We present RLFactory, a plug-and-play reinforcement learning post-training framework for multi-round to…

Reinforcement LearningNatural Questions

An Architecture for Deploying Reinforcement Learning in Industrial Environments

2023-06-02 · Georg Schäfer, Reuf Kozlica, Stefan Wegenkittl, Stefan Huber

Industry 4.0 is driven by demands like shorter time-to-market, mass customization of products, and batch size one production. Reinforcement Learning (RL), a machine learning paradigm shown to possess a great potential in…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Plug-Tagger: A Pluggable Sequence Labeling Framework Using Language Models

2021-10-14 · Xin Zhou, Ruotian Ma, Tao Gui, Yiding Tan 외

Plug-and-play functionality allows deep learning models to adapt well to different tasks without requiring any parameters modified. Recently, prefix-tuning was shown to be a plug-and-play method on various text generatio…

Language ModellingText Generation

AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

2023-08-07 · Michaël Mathieu, Sherjil Ozair, Srivatsan Srinivasan, Caglar Gulcehre 외

StarCraft II is one of the most challenging simulated reinforcement learning environments; it is partially observable, stochastic, multi-agent, and mastering StarCraft II requires strategic planning over long time horizo…

Offline RLreinforcement-learningReinforcement LearningStarcraft+1