paper-with-me

Papers

Regret Bounds for Robust Adaptive Control of the Linear Quadratic Regulator

2018-05-23 · NeurIPS 2018 12 · Sarah Dean, Horia Mania, Nikolai Matni, Benjamin Recht, Stephen Tu

We consider adaptive control of the Linear Quadratic Regulator (LQR), where an unknown linear system is controlled subject to quadratic costs. Leveraging recent developments in the estimation of linear systems and in robust controller synthesis, we present the first provably polynomial time algorithm that provides high probability guarantees of sub-linear regret on this problem. We further study the interplay between regret minimization and parameter estimation by proving a lower bound on the expected regret in terms of the exploration schedule used by any algorithm. Finally, we conduct a numerical study comparing our robust adaptive algorithm to other methods from the adaptive LQR literature, and demonstrate the flexibility of our proposed method by extending it to a demand forecasting problem subject to state constraints.

📄 PDF Abstract BibTeX arXiv:1805.09388

Code (0)

등록된 구현이 없습니다.

Tasks

Demand Forecastingparameter estimation

Similar Papers 제목 키워드 기반

Regret Bounds for Episodic Risk-Sensitive Linear Quadratic Regulator

2024-06-08 · Wenhao Xu, Xuefeng Gao, Xuedong He

Risk-sensitive linear quadratic regulator is one of the most fundamental problems in risk-sensitive optimal control. In this paper, we study online adaptive control of risk-sensitive linear quadratic regulator in the fin…

Regret Lower Bounds for Learning Linear Quadratic Gaussian Systems

2022-01-05 · Ingvar Ziemann, Henrik Sandberg

TWe establish regret lower bounds for adaptively controlling an unknown linear Gaussian system with quadratic costs. We combine ideas from experiment design, estimation theory and a perturbation bound of certain informat…

Towards a Dimension-Free Understanding of Adaptive Linear Control

2021-03-19 · Juan C. Perdomo, Max Simchowitz, Alekh Agarwal, Peter Bartlett

We study the problem of adaptive control of the linear quadratic regulator for systems in very high, or even infinite dimension. We demonstrate that while sublinear regret requires finite dimensional inputs, the ambient …

Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator

2026-04-24 · Spencer Hutchinson, Nanfei Jiang, Mahnoosh Alizadeh arxiv

We study the problem of adaptive control of the stochastic linear quadratic regulator (LQR) with constraints that must be satisfied at every time step. Prior work on the multidimensional problem has shown $\tilde{O}(T^{2…

Nonasymptotic Regret Analysis of Adaptive Linear Quadratic Control with Model Misspecification

2023-12-29 · Bruce D. Lee, Anders Rantzer, Nikolai Matni

The strategy of pre-training a large model on a diverse dataset, then fine-tuning for a particular application has yielded impressive results in computer vision, natural language processing, and robotic control. This str…