paper-with-me

홈 › Papers

Regret Analysis of Learning-Based MPC with Partially-Unknown Cost Function

2021-08-04 · Ilgin Dogan, Zuo-Jun Max Shen, Anil Aswani

The exploration/exploitation trade-off is an inherent challenge in data-driven adaptive control. Though this trade-off has been studied for multi-armed bandits (MAB's) and reinforcement learning for linear systems; it is less well-studied for learning-based control of nonlinear systems. A significant theoretical challenge in the nonlinear setting is that there is no explicit characterization of an optimal controller for a given set of cost and system parameters. We propose the use of a finite-horizon oracle controller with full knowledge of parameters as a reasonable surrogate to optimal controller. This allows us to develop policies in the context of learning-based MPC and MAB's and conduct a control-theoretic analysis using techniques from MPC- and optimization-theory to show these policies achieve low regret with respect to this finite-horizon oracle. Our simulations exhibit the low regret of our policy on a heating, ventilation, and air-conditioning model with partially-unknown cost function.

📄 PDF Abstract BibTeX arXiv:2108.02307

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Armed Bandits

Similar Papers 제목 키워드 기반

To Explore or Not to Explore: Regret-Based LTL Planning in Partially-Known Environments

2022-04-01 · Jianing Zhao, Keyi Zhu, Mingyang Feng, Xiang Yin

In this paper, we investigate the optimal robot path planning problem for high-level specifications described by co-safe linear temporal logic (LTL) formulae. We consider the scenario where the map geometry of the worksp…

Regret Analysis of Distributed Online LQR Control for Unknown LTI Systems

2021-05-15 · Ting-Jui Chang, Shahin Shahrampour

Online optimization has recently opened avenues to study optimal control for time-varying cost functions that are unknown in advance. Inspired by this line of research, we study the distributed online linear quadratic re…

Regret Analysis of Distributed Online Control for LTI Systems with Adversarial Disturbances

2023-10-04 · Ting-Jui Chang, Shahin Shahrampour

This paper addresses the distributed online control problem over a network of linear time-invariant (LTI) systems (with possibly unknown dynamics) in the presence of adversarial perturbations. There exists a global netwo…

Online Learning for Unknown Partially Observable MDPs

2021-02-25 · Mehdi Jafarnia-Jahromi, Rahul Jain, Ashutosh Nayyar

Solving Partially Observable Markov Decision Processes (POMDPs) is hard. Learning optimal controllers for POMDPs when the model is unknown is harder. Online learning of optimal controllers for unknown POMDPs, which requi…

Finite-time System Identification and Adaptive Control in Autoregressive Exogenous Systems

2021-08-26 · Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima Anandkumar

Autoregressive exogenous (ARX) systems are the general class of input-output dynamical systems used for modeling stochastic linear dynamical systems (LDS) including partially observable LDS such as LQG systems. In this w…