paper-with-me

홈 › Papers

Regret Analysis of Learning-Based Linear Quadratic Gaussian Control with Additive Exploration

2023-11-05 · Archith Athrey, Othmane Mazhar, Meichen Guo, Bart De Schutter, Shengling Shi

In this paper, we analyze the regret incurred by a computationally efficient exploration strategy, known as naive exploration, for controlling unknown partially observable systems within the Linear Quadratic Gaussian (LQG) framework. We introduce a two-phase control algorithm called LQG-NAIVE, which involves an initial phase of injecting Gaussian input signals to obtain a system model, followed by a second phase of an interplay between naive exploration and control in an episodic fashion. We show that LQG-NAIVE achieves a regret growth rate of $\tilde{\mathcal{O}}(\sqrt{T})$, i.e., $\mathcal{O}(\sqrt{T})$ up to logarithmic factors after $T$ time steps, and we validate its performance through numerical simulations. Additionally, we propose LQG-IF2E, which extends the exploration signal to a `closed-loop' setting by incorporating the Fisher Information Matrix (FIM). We provide compelling numerical evidence of the competitive performance of LQG-IF2E compared to LQG-NAIVE.

📄 PDF Abstract BibTeX arXiv:2311.02679

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient Exploration

Similar Papers 제목 키워드 기반

Adaptive Control and Regret Minimization in Linear Quadratic Gaussian (LQG) Setting

2020-03-12 · Sahin Lale, Kamyar Azizzadenesheli, Babak Hassibi, Anima Anandkumar

We study the problem of adaptive control in partially observable linear quadratic Gaussian control systems, where the model dynamics are unknown a priori. We propose LqgOpt, a novel reinforcement learning algorithm based…

Reinforcement Learning

Regret Lower Bounds for Learning Linear Quadratic Gaussian Systems

2022-01-05 · Ingvar Ziemann, Henrik Sandberg

TWe establish regret lower bounds for adaptively controlling an unknown linear Gaussian system with quadratic costs. We combine ideas from experiment design, estimation theory and a perturbation bound of certain informat…

Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG

2024-03-13 · Ting-Jui Chang, Shahin Shahrampour

Recent advancement in online optimization and control has provided novel tools to study online linear quadratic regulator (LQR) problems, where cost matrices are varying adversarially over time. However, the controller p…

Regret Bounds for Episodic Risk-Sensitive Linear Quadratic Regulator

2024-06-08 · Wenhao Xu, Xuefeng Gao, Xuedong He

Risk-sensitive linear quadratic regulator is one of the most fundamental problems in risk-sensitive optimal control. In this paper, we study online adaptive control of risk-sensitive linear quadratic regulator in the fin…

Riccati updates for online linear quadratic control

2020-06-08 · L4DC 2020 6 · Mohammad Akbari, Bahman Gharesifard, Tamas Linder

We study an online setting of the linear quadratic Gaussian optimal control problem on a sequence of cost functions, where similar to classical online optimization, the future decisions are made by only knowing the cost …