paper-with-me

홈 › Papers

An Online Model-Following Projection Mechanism Using Reinforcement Learning

2023-02-05 · Mohammed I. Abouheaf, Hashim A. Hashim, Mohammad A. Mayyas, Kyriakos G. Vamvoudakis

In this paper, we propose a model-free adaptive learning solution for a model-following control problem. This approach employs policy iteration, to find an optimal adaptive control solution. It utilizes a moving finite-horizon of model-following error measurements. In addition, the control strategy is designed by using a projection mechanism that employs Lagrange dynamics. It allows for real-time tuning of derived actor-critic structures to find the optimal model-following strategy and sustain optimized adaptation performance. Finally, the efficacy of the proposed framework is emphasized through a comparison with sliding mode and high-order model-free adaptive control approaches. Keywords: Model Reference Adaptive Systems, Reinforcement Learning, adaptive critics, control system, stochastic, nonlinear system

📄 PDF Abstract BibTeX arXiv:2302.02493

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

An Observer-Based Reinforcement Learning Solution for Model-Following Problems

2023-08-19 · Mohammed I. Abouheaf, Kyriakos G. Vamvoudakis, Mohammad A. Mayyas, Hashim A. Hashim

In this paper, a multi-objective model-following control problem is solved using an observer-based adaptive learning scheme. The overall goal is to regulate the model-following error dynamics along with optimizing the dy…

reinforcement-learningReinforcement Learning

Provably Safe Reinforcement Learning via Action Projection using Reachability Analysis and Polynomial Zonotopes

2022-10-19 · Niklas Kochdumper, Hanna Krasowski, Xiao Wang, Stanley Bak 외

While reinforcement learning produces very promising results for many applications, its main disadvantage is the lack of safety guarantees, which prevents its use in safety-critical systems. In this work, we address this…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Safe Reinforcement Learning

The Real, the Better: Aligning Large Language Models with Online Human Behaviors

2024-05-01 · Guanying Jiang, Lingyong Yan, Haibo Shi, Dawei Yin

Large language model alignment is widely used and studied to avoid LLM producing unhelpful and harmful responses. However, the lengthy training process and predefined preference bias hinder adaptation to online diverse h…

Language ModelingLanguage ModellingLarge Language Model

From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning

2026-05-13 · Timofey Tomashevskiy arxiv

Safety in reinforcement learning is often specified through cumulative cost constraints, but these trajectory-level guarantees do not directly prevent unsafe individual decisions, especially under nonstationarity. In con…

Reinforcement Learning

Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms

2026-03-15 · Jingyi Liu, Jian Guo, Eberhard Gill arxiv

Reinforcement learning has proven its power on various occasions. However, its performance is not always guaranteed when system dynamics change. Instead, it largely relies on users' empirical experience. For reinforcemen…

Reinforcement Learning