paper-with-me

Papers

Steady-State Error Compensation in Reference Tracking and Disturbance Rejection Problems for Reinforcement Learning-Based Control

2022-01-31 · Daniel Weber, Maximilian Schenke, Oliver Wallscheid

Reinforcement learning (RL) is a promising, upcoming topic in automatic control applications. Where classical control approaches require a priori system knowledge, data-driven control approaches like RL allow a model-free controller design procedure, rendering them emergent techniques for systems with changing plant structures and varying parameters. While it was already shown in various applications that the transient control behavior for complex systems can be sufficiently handled by RL, the challenge of non-vanishing steady-state control errors remains, which arises from the usage of control policy approximations and finite training times. To overcome this issue, an integral action state augmentation (IASA) for actor-critic-based RL controllers is introduced that mimics an integrating feedback, which is inspired by the delta-input formulation within model predictive control. This augmentation does not require any expert knowledge, leaving the approach model free. As a result, the RL controller learns how to suppress steady-state control deviations much more effectively. Two exemplary applications from the domain of electrical energy engineering validate the benefit of the developed method both for reference tracking and disturbance rejection. In comparison to a standard deep deterministic policy gradient (DDPG) setup, the suggested IASA extension allows to reduce the steady-state error by up to 52 $\%$ within the considered validation scenarios.

📄 PDF Abstract BibTeX arXiv:2201.13331

Code (1)

webbah/sec-for-reinforcement-learning 공식 구현

Tasks

Model Predictive ControlReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Zero-Error Tracking for Autonomous Vehicles through Epsilon-Trajectory Generation

2020-07-20

This paper presents a control method and trajectory planner for vehicles with first-order nonholonomic constraints that guarantee asymptotic convergence to a time-indexed trajectory. To overcome the nonholonomic constrai…

Autonomous Vehicles

Design and Performance Analysis of a Class of Generalized Predictive Controllers

2023-11-06 · Feilong Zhang

The design and structure of generalized predictive control (GPC) are not simple and intuitive. The performance analysis does not deeply analyze how the controller parameters affect the system characteristics and the rela…

Tracking performance of PID for nonlinear stochastic systems

2023-03-19 · Cheng Zhao, Shuo Yuan

In this paper, we will consider a class of continuous-time stochastic control systems with both unknown nonlinear structure and unknown disturbances, and investigate the capability of the classical proportional-integral-…

Towards Dynamic Model Identification and Gravity Compensation for the dVRK-Si Patient Side Manipulator

2026-03-12 · Haoying Zhou, Hao Yang, Brendan Burkhart, Anton Deguet 외 arxiv

The da Vinci Research Kit (dVRK) is widely used for research in robot-assisted surgery, but most modeling and control methods target the first-generation dVRK Classic. The recently introduced dVRK-Si, built from da Vinci…

Learning disturbance models for offset-free reference tracking

2023-12-18 · Pablo Krupa, Mario Zanon, Alberto Bemporad

This work presents a nonlinear MPC framework that guarantees asymptotic offset-free tracking of generic reference trajectories by learning a nonlinear disturbance model, which compensates for input disturbances and model…