paper-with-me

홈 › Papers

Optimism-Based Adaptive Regulation of Linear-Quadratic Systems

2017-11-20 · Mohamad Kazem Shirani Faradonbeh, Ambuj Tewari, George Michailidis

The main challenge for adaptive regulation of linear-quadratic systems is the trade-off between identification and control. An adaptive policy needs to address both the estimation of unknown dynamics parameters (exploration), as well as the regulation of the underlying system (exploitation). To this end, optimism-based methods which bias the identification in favor of optimistic approximations of the true parameter are employed in the literature. A number of asymptotic results have been established, but their finite time counterparts are few, with important restrictions. This study establishes results for the worst-case regret of optimism-based adaptive policies. The presented high probability upper bounds are optimal up to logarithmic factors. The non-asymptotic analysis of this work requires very mild assumptions; (i) stabilizability of the system's dynamics, and (ii) limiting the degree of heaviness of the noise distribution. To establish such bounds, certain novel techniques are developed to comprehensively address the probabilistic behavior of dependent random matrices with heavy-tailed distributions.

📄 PDF Abstract BibTeX arXiv:1711.07230

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Adaptive Control and Regret Minimization in Linear Quadratic Gaussian (LQG) Setting

2020-03-12 · Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima Anandkumar

We study the problem of adaptive control in partially observable linear quadratic Gaussian control systems, where the model dynamics are unknown a priori. We propose LqgOpt, a novel reinforcement learning algorithm based…

Reinforcement Learning

Regret Minimization in Partially Observable Linear Quadratic Control

2020-01-31 · Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima Anandkumar

We study the problem of regret minimization in partially observable linear quadratic control systems when the model dynamics are unknown a priori. We propose ExpCommit, an explore-then-commit algorithm that learns the mo…

Finite-time System Identification and Adaptive Control in Autoregressive Exogenous Systems

2021-08-26 · Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima Anandkumar

Autoregressive exogenous (ARX) systems are the general class of input-output dynamical systems used for modeling stochastic linear dynamical systems (LDS) including partially observable LDS such as LQG systems. In this w…

Exponentially Stable Adaptive Optimal Control of Uncertain LTI Systems

2022-05-05 · Anton Glushchenko, Konstantin Lastochkin

A novel method of an adaptive linear quadratic (LQ) regulation of uncertain continuous linear time-invariant systems is proposed. Such an approach is based on the direct self-tuning regulators design framework and the ex…

Augmented RBMLE-UCB Approach for Adaptive Control of Linear Quadratic Systems

2022-01-25 · Akshay Mete, Rahul Singh, P. R. Kumar

We consider the problem of controlling an unknown stochastic linear system with quadratic costs - called the adaptive LQ control problem. We re-examine an approach called ''Reward Biased Maximum Likelihood Estimate'' (RB…

parameter estimationThompson Sampling