paper-with-me

Papers

Reinforcement learning for linear-convex models with jumps via stability analysis of feedback controls

2021-04-19 · Xin Guo, Anran Hu, Yufei Zhang

We study finite-time horizon continuous-time linear-convex reinforcement learning problems in an episodic setting. In this problem, the unknown linear jump-diffusion process is controlled subject to nonsmooth convex costs. We show that the associated linear-convex control problems admit Lipchitz continuous optimal feedback controls and further prove the Lipschitz stability of the feedback controls, i.e., the performance gap between applying feedback controls for an incorrect model and for the true model depends Lipschitz-continuously on the magnitude of perturbations in the model coefficients; the proof relies on a stability analysis of the associated forward-backward stochastic differential equation. We then propose a novel least-squares algorithm which achieves a regret of the order $O(\sqrt{N\ln N})$ on linear-convex learning problems with jumps, where $N$ is the number of learning episodes; the analysis leverages the Lipschitz stability of feedback controls and concentration properties of sub-Weibull random variables. Numerical experiment confirms the convergence and the robustness of the proposed algorithm.

📄 PDF Abstract BibTeX arXiv:2104.09311

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Stability Analysis and State-Feedback Stabilization of LPV Time-Delay Systems with Piecewise Constant Parameters subject to Spontaneous Poissonian Jumps

2021-02-09 · Muhammad Zakwan

This paper discusses the stability analysis of linear parameter varying systems with a parameter-dependent delay where the parameters are assumed to be stochastic piecewise constants under spontaneous Poissonian jumps. B…

Incremental Stability and Performance Analysis of Discrete-Time Nonlinear Systems using the LPV Framework

2021-03-19 · Patrick J. W. Koelewijn, Roland Tóth

The dissipativity framework is widely used to analyze stability and performance of nonlinear systems. By embedding nonlinear systems in an LPV representation, the convex tools of the LPV framework can be applied to nonli…

Convex Equilibrium-Free Stability and Performance Analysis of Discrete-Time Nonlinear Systems

2024-02-15 · Patrick J. W. Koelewijn, Siep Weiland, Roland Tóth

This paper considers the equilibrium-free stability and performance analysis of discrete-time nonlinear systems. We consider two types of equilibrium-free notions. Namely, the universal shifted concept, which considers s…

Fine-grained Analysis of Stability and Generalization for Stochastic Bilevel Optimization

2026-04-05 · Xuelin Zhang, Hong Chen, Bin Gu, Tieliang Gong 외 arxiv

Stochastic bilevel optimization (SBO) has been integrated into many machine learning paradigms recently, including hyperparameter optimization, meta learning, and reinforcement learning. Along with the wide range of appl…

Hyperparameter OptimizationReinforcement LearningBilevel Optimization

A hybrid systems framework for data-based adaptive control of linear time-varying systems

2024-05-23 · Andrea Iannelli, Romain Postoyan

We consider the data-driven stabilization of discrete-time linear time-varying systems. The controller is defined as a linear state-feedback law whose gain is adapted to the plant changes through a data-based event-trigg…