paper-with-me

홈 › Papers

Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels

2024-01-02 · Lintao Ye, Aritra Mitra, Vijay Gupta

Consider a linear quadratic regulator (LQR) problem being solved in a model-free manner using the policy gradient approach. If the gradient of the quadratic cost is being transmitted across a rate-limited channel, both the convergence and the rate of convergence of the resulting controller may be affected by the bit-rate permitted by the channel. We first pose this problem in a communication-constrained optimization framework and propose a new adaptive quantization algorithm titled Adaptively Quantized Gradient Descent (AQGD). This algorithm guarantees exponentially fast convergence to the globally optimal policy, with no deterioration of the exponent relative to the unquantized setting, above a certain finite threshold bit-rate allowed by the communication channel. We then propose a variant of AQGD that provides similar performance guarantees when applied to solve the model-free LQR problem. Our approach reveals the benefits of adaptive quantization in preserving fast linear convergence rates, and, as such, may be of independent interest to the literature on compressed optimization. Our work also marks a first step towards a more general bridge between the fields of model-free control design and networked control systems.

📄 PDF Abstract BibTeX arXiv:2401.01258

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Similar Papers 제목 키워드 기반

Online Policy Gradient for Model Free Learning of Linear Quadratic Regulators with $\sqrt{T}$ Regret

2021-02-25 · Asaf Cassel, Tomer Koren

We consider the task of learning to control a linear dynamical system under fixed quadratic costs, known as the Linear Quadratic Regulator (LQR) problem. While model-free approaches are often favorable in practice, thus …

Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR

2024-01-25 · Leonardo F. Toso, Donglin Zhan, James Anderson, Han Wang

We investigate the problem of learning linear quadratic regulators (LQR) in a multi-task, heterogeneous, and model-free setting. We characterize the stability and personalization guarantees of a policy gradient-based (PG…

Meta-Learning

Linear-Quadratic regulators for internal boundary control of lane-free automated vehicle traffic

2020-12-31 · Milad Malekzadeh, Ioannis Papamichail, Markos Papageorgiou

Lane-free vehicle movement has been recently proposed for connected automated vehicles (CAV) due to various potential advantages. One such advantage stems from the fact that incremental changes of the road width in lane-…

Learning the model-free linear quadratic regulator via random search

2020-06-08 · L4DC 2020 6 · Hesameddin Mohammadi, Mihailo R. Jovanovic', Mahdi Soltanolkotabi

Model-free reinforcement learning attempts to find an optimal control action for an unknown dynamical system by directly searching over the parameter space of controllers. The convergence behavior and statistical propert…

Reinforcement Learning (RL)

Regret Bounds for Episodic Risk-Sensitive Linear Quadratic Regulator

2024-06-08 · Wenhao Xu, Xuefeng Gao, Xuedong He

Risk-sensitive linear quadratic regulator is one of the most fundamental problems in risk-sensitive optimal control. In this paper, we study online adaptive control of risk-sensitive linear quadratic regulator in the fin…