paper-with-me

Papers

Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters

2023-03-29 · Deyue Li

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. In this general setting, we apply the policy gradient method, a reinforcement learning technique, to search for the optimal control without requiring knowledge of statistical information of the parameters. We investigate the sub-Gaussianity of the state process and establish global linear convergence guarantee for this approach based on assumptions that are weaker and easier to verify compared to existing results. Numerical experiments are presented to illustrate our result.

📄 PDF Abstract BibTeX arXiv:2303.16548

Code (0)

등록된 구현이 없습니다.

Tasks

Policy Gradient Methodsreinforcement-learning

Similar Papers 제목 키워드 기반

Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

2022-11-01 · Michael Giegrich, Christoph Reisinger, Yufei Zhang

We study the global linear convergence of policy gradient (PG) methods for finite-horizon continuous-time exploratory linear-quadratic control (LQC) problems. The setting includes stochastic LQC problems with indefinite …

Policy Gradient Methods

Optimization Landscape of Policy Gradient Methods for Discrete-time Static Output Feedback

2023-10-29 · Jingliang Duan, Jie Li, Xuyang Chen, Kai Zhao 외

In recent times, significant advancements have been made in delving into the optimization landscape of policy gradient methods for achieving optimal control in linear time-invariant (LTI) systems. Compared with state-fee…

Policy Gradient Methods

Global Convergence Using Policy Gradient Methods for Model-free Markovian Jump Linear Quadratic Control

2021-11-30 · Santanu Rathod, Manoj Bhadu, Abir De

Owing to the growth of interest in Reinforcement Learning in the last few years, gradient based policy control methods have been gaining popularity for Control problems as well. And rightly so, since gradient policy meth…

Policy Gradient Methods

Policy Optimization for Markovian Jump Linear Quadratic Control: Gradient-Based Methods and Global Convergence

2020-11-24 · Joao Paulo Jansch-Porto, Bin Hu, Geir Dullerud

Recently, policy optimization for control purposes has received renewed attention due to the increasing interest in reinforcement learning. In this paper, we investigate the global convergence of gradient-based policy op…

Policy Gradient Methods

Global Convergence of Policy Gradient for Linear-Quadratic Mean-Field Control/Game in Continuous Time

2020-08-16 · Weichen Wang, Jiequn Han, Zhuoran Yang, Zhaoran Wang

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a me…