paper-with-me

홈 › Papers

Regularization for Covariance Parameterization of Direct Data-Driven LQR Control

2025-03-04 · Feiran Zhao, Alessandro Chiuso, Florian Dörfler

As the benchmark of data-driven control methods, the linear quadratic regulator (LQR) problem has gained significant attention. A growing trend is direct LQR design, which finds the optimal LQR gain directly from raw data and bypassing system identification. To achieve this, our previous work develops a direct LQR formulation parameterized by sample covariance. In this paper, we propose a regularization method for the covariance-parameterized LQR. We show that the regularizer accounts for the uncertainty in both the steady-state covariance matrix corresponding to closed-loop stability, and the LQR cost function corresponding to averaged control performance. With a positive or negative coefficient, the regularizer can be interpreted as promoting either exploitation or exploration, which are well-known trade-offs in reinforcement learning. In simulations, we observe that our covariance-parameterized LQR with regularization can significantly outperform the certainty-equivalence LQR in terms of both the optimality gap and the robust closed-loop stability.

📄 PDF Abstract BibTeX arXiv:2503.02985

Code (1)

feiran-zhao-eth/policy-gradient-adaptive-control

Similar Papers 제목 키워드 기반

A Comparative Theoretical Analysis of Entropy Control Methods in Reinforcement Learning

2026-04-02 · Ming Lei, Christophe Baehr arxiv

Reinforcement learning (RL) has become a key approach for enhancing reasoning in large language models (LLMs), yet scalable training is often hindered by the rapid collapse of policy entropy, which leads to premature con…

Reinforcement Learning

Bias correction and instrumental variables for direct data-driven model-reference control

2024-11-08 · Manas Mejari, Valentina Breschi, Simone Formentin, Dario Piga

Managing noisy data is a central challenge in direct data-driven control design. We propose an approach for synthesizing model-reference controllers for linear time-invariant (LTI) systems using noisy state-input data, e…

Asymptotics of Linear Regression with Linearly Dependent Data

2024-12-04 · Behrad Moniri, Hamed Hassani

In this paper we study the asymptotics of linear regression in settings with non-Gaussian covariates where the covariates exhibit a linear dependency structure, departing from the standard assumption of independence. We …

regression

Linear Convergence of Data-Enabled Policy Optimization for Linear Quadratic Tracking

2024-10-08 · Shubo Kang, Feiran Zhao, Keyou You

Data-enabled policy optimization (DeePO) is a newly proposed method to attack the open problem of direct adaptive LQR. In this work, we extend the DeePO framework to the linear quadratic tracking (LQT) with offline data.…

Data-Aided Regularization of Direct-Estimate Combiner in Distributed MIMO Systems

2025-01-21 · Bikshapathi Gouda, Italo Atzeni, Antti Tölli

This paper explores the data-aided regularization of the direct-estimate combiner in the uplink of a distributed multiple-input multiple-output system. The network-wide combiner can be computed directly from the pilot si…