paper-with-me

홈 › Papers

Learning the model-free linear quadratic regulator via random search

2020-06-08 · L4DC 2020 6 · Hesameddin Mohammadi, Mihailo R. Jovanovic', Mahdi Soltanolkotabi

Model-free reinforcement learning attempts to find an optimal control action for an unknown dynamical system by directly searching over the parameter space of controllers. The convergence behavior and statistical properties of these approaches are often poorly understood because of the nonconvex nature of the underlying optimization problems as well as the lack of exact gradient computation. In this paper, we examine the standard infinite-horizon linear quadratic regulator problem for continuous-time systems with unknown state-space parameters. We provide theoretical bounds on the convergence rate and sample complexity of a random search method. Our results demonstrate that the required simulation time for achieving $\epsilon$-accuracy in a model-free setup and the total number of function evaluations are both of $O (\log \, (1/\epsilon) )$.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Convergence and sample complexity of gradient methods for the model-free linear quadratic regulator problem

2019-12-26 · Hesameddin Mohammadi, Armin Zare, Mahdi Soltanolkotabi, Mihailo R. Jovanović

Model-free reinforcement learning attempts to find an optimal control action for an unknown dynamical system by directly searching over the parameter space of controllers. The convergence behavior and statistical propert…

Reinforcement Learning

Online Policy Gradient for Model Free Learning of Linear Quadratic Regulators with $\sqrt{T}$ Regret

2021-02-25 · Asaf Cassel, Tomer Koren

We consider the task of learning to control a linear dynamical system under fixed quadratic costs, known as the Linear Quadratic Regulator (LQR) problem. While model-free approaches are often favorable in practice, thus …

Simple random search provides a competitive approach to reinforcement learning

2018-03-19 · Horia Mania, Aurelia Guy, Benjamin Recht

A common belief in model-free reinforcement learning is that methods based on random search in the parameter space of policies exhibit significantly worse sample complexity than those that explore the space of actions. W…

Computational Efficiencycontinuous-controlContinuous ControlMuJoCo+3

Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR

2024-01-25 · Leonardo F. Toso, Donglin Zhan, James Anderson, Han Wang

We investigate the problem of learning linear quadratic regulators (LQR) in a multi-task, heterogeneous, and model-free setting. We characterize the stability and personalization guarantees of a policy gradient-based (PG…

Meta-Learning

Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters

2023-03-29 · Deyue Li

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. I…

Policy Gradient Methodsreinforcement-learning