paper-with-me

홈 › Papers

Combining Model-Based and Model-Free Methods for Nonlinear Control: A Provably Convergent Policy Gradient Approach

2020-06-12 · Guannan Qu, Chenkai Yu, Steven Low, Adam Wierman

Model-free learning-based control methods have seen great success recently. However, such methods typically suffer from poor sample complexity and limited convergence guarantees. This is in sharp contrast to classical model-based control, which has a rich theory but typically requires strong modeling assumptions. In this paper, we combine the two approaches to achieve the best of both worlds. We consider a dynamical system with both linear and non-linear components and develop a novel approach to use the linear model to define a warm start for a model-free, policy gradient method. We show this hybrid approach outperforms the model-based controller while avoiding the convergence issues associated with model-free approaches via both numerical experiments and theoretical analyses, in which we derive sufficient conditions on the non-linear component such that our approach is guaranteed to converge to the (nearly) global optimal controller.

📄 PDF Abstract BibTeX arXiv:2006.07476

Code (0)

등록된 구현이 없습니다.

Tasks

model

Similar Papers 제목 키워드 기반

Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems

2024-03-06 · Wesley A. Suttle, Vipul K. Sharma, Krishna C. Kosaraju, S. Sivaranjani 외

We develop provably safe and convergent reinforcement learning (RL) algorithms for control of nonlinear dynamical systems, bridging the gap between the hard safety guarantees of control theory and the convergence guarant…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning

Sample-efficient Safe Learning for Online Nonlinear Control with Control Barrier Functions

2022-07-29 · Wenhao Luo, Wen Sun, Ashish Kapoor

Reinforcement Learning (RL) and continuous nonlinear control have been successfully deployed in multiple domains of complicated sequential decision-making tasks. However, given the exploration nature of the learning proc…

Decision MakingReinforcement Learning (RL)Safe ExplorationSequential Decision Making

Provably-Stable Neural Network-Based Control of Nonlinear Systems

2025-02-01 · Anran Li, John P. Swensen, Mehdi Hosseinzadeh

In recent years, Neural Networks (NNs) have been employed to control nonlinear systems due to their potential capability in dealing with situations that might be difficult for conventional nonlinear control schemes. Howe…

Neural Lyapunov Control

2020-05-01 · NeurIPS 2019 12 · Ya-Chien Chang, Nima Roohi, Sicun Gao

We propose new methods for learning control policies and neural network Lyapunov functions for nonlinear control problems, with provable guarantee of stability. The framework consists of a learner that attempts to find t…

High-Dimensional Random Projection for Activation Steering in Language Models

2026-06-13 · Minh-Hieu Pham, Bach Do, Laziz Abdullaev, Tan Minh Nguyen 외 arxiv

Activation steering has emerged as a key methodology for controlling the behavior of large language models (LLMs). Existing difference-in-means based methods, however, are fundamentally limited: they capture only mean di…