paper-with-me

홈 › Papers

Deep Value Model Predictive Control

2019-10-08 · Farbod Farshidian, David Hoeller, Marco Hutter

In this paper, we introduce an actor-critic algorithm called Deep Value Model Predictive Control (DMPC), which combines model-based trajectory optimization with value function estimation. The DMPC actor is a Model Predictive Control (MPC) optimizer with an objective function defined in terms of a value function estimated by the critic. We show that our MPC actor is an importance sampler, which minimizes an upper bound of the cross-entropy to the state distribution of the optimal sampling policy. In our experiments with a Ballbot system, we show that our algorithm can work with sparse and binary reward signals to efficiently solve obstacle avoidance and target reaching tasks. Compared to previous work, we show that including the value function in the running cost of the trajectory optimizer speeds up the convergence. We also discuss the necessary strategies to robustify the algorithm in practice.

📄 PDF Abstract BibTeX arXiv:1910.03358

Code (0)

등록된 구현이 없습니다.

Tasks

modelModel Predictive Control

Similar Papers 제목 키워드 기반

Predictive Control with Learning-Based Terminal Costs Using Approximate Value Iteration

2022-12-01 · Francisco Moreno-Mora, Lukas Beckenbach, Stefan Streif

Stability under model predictive control (MPC) schemes is frequently ensured by terminal ingredients. Employing a (control) Lyapunov function as the terminal cost constitutes a common choice. Learning-based methods may b…

Model Predictive Control

Soft MPCritic: Amortized Model Predictive Value Iteration

2026-04-01 · Thomas Banker, Nathan P. Lawrence, Ali Mesbah arxiv

Reinforcement learning (RL) and model predictive control (MPC) offer complementary strengths, yet combining them at scale remains computationally challenging. We propose soft MPCritic, an RL-MPC framework that learns in …

Reinforcement Learning

Bootstrapped Model Predictive Control

2025-03-24 · Yuhang Wang, Hanwei Guo, Sizhe Wang, Long Qian 외

Model Predictive Control (MPC) has been demonstrated to be effective in continuous control tasks. When a world model and a value function are available, planning a sequence of actions ahead of time leads to a better poli…

continuous-controlContinuous ControlImitation Learningmodel+1

Lessons from AlphaZero for Optimal, Model Predictive, and Adaptive Control

2021-08-20 · Dimitri Bertsekas

In this paper we aim to provide analysis and insights (often based on visualization), which explain the beneficial effects of on-line decision making on top of off-line training. In particular, through a unifying abstrac…

Bayesian OptimizationDecision MakingModel Predictive Control

Efficient Recursive Data-enabled Predictive Control (Extended Version)

2023-09-24 · Jicheng Shi, Yingzhao Lian, Colin N. Jones

In the field of model predictive control, Data-enabled Predictive Control (DeePC) offers direct predictive control, bypassing traditional modeling. However, challenges emerge with increased computational demand due to re…

FormModel Predictive Control