MRAC-RL: A Framework for On-Line Policy Adaptation Under Parametric Model Uncertainty
Reinforcement learning (RL) algorithms have been successfully used to develop control policies for dynamical systems. For many such systems, these policies are trained in a simulated environment. Due to discrepancies between the simulated model and the true system dynamics, RL trained policies often fail to generalize and adapt appropriately when deployed in the real-world environment. Current research in bridging this sim-to-real gap has largely focused on improvements in simulation design and on the development of improved and specialized RL algorithms for robust control policy generation. In this paper we apply principles from adaptive control and system identification to develop the model-reference adaptive control & reinforcement learning (MRAC-RL) framework. We propose a set of novel MRAC algorithms applicable to a broad range of linear and nonlinear systems, and derive the associated control laws. The MRAC-RL framework utilizes an inner-loop adaptive controller that allows a simulation-trained outer-loop policy to adapt and operate effectively in a test environment, even when parametric model uncertainty exists. We demonstrate that the MRAC-RL approach improves upon state-of-the-art RL algorithms in developing control policies that can be applied to systems with modeling errors.
Code (0)
등록된 구현이 없습니다.
Tasks
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Online Policies for Real-Time Control Using MRAC-RL
In this paper, we propose the Model Reference Adaptive Control & Reinforcement Learning (MRAC-RL) approach to developing online policies for systems in which modeling errors occur in real-time. Although reinforcement lea…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)I-DREM MRAC with Time-Varying Adaptation Rate & No A Priori Knowledge of Control Input Matrix Sign to Relax PE Condition
The known dynamic regressor extension and mixing method (DREM) is combined with the proposed filter of a new type, which uses the integration operation with forgetting, and the recursive least-squares method to develop t…
Arbitrarily Fast Multivariable Least-squares MRAC
A novel least-squares model-reference direct adaptive control (LS-MRAC) algorithm for multivariable (MIMO) plants is presented. The controller parameters are directly updated based on the output tracking error. The contr…
Deep Model Reference Adaptive Control
We present a new neuroadaptive architecture: Deep Neural Network based Model Reference Adaptive Control (DMRAC). Our architecture utilizes the power of deep neural network representations for modeling significant nonline…
modelAdaptive Output Tracking Control with Reference Model System Uncertainties
This paper develops adaptive output tracking control schemes with the reference output signal generated from an unknown reference system whose output derivatives are also unknown. To deal with such reference system uncer…