paper-with-me

홈 › Papers

Online Observer-Based Inverse Reinforcement Learning

2020-11-03 · Ryan Self, Kevin Coleman, He Bai, Rushikesh Kamalapurkar

In this paper, a novel approach to the output-feedback inverse reinforcement learning (IRL) problem is developed by casting the IRL problem, for linear systems with quadratic cost functions, as a state estimation problem. Two observer-based techniques for IRL are developed, including a novel observer method that re-uses previous state estimates via history stacks. Theoretical guarantees for convergence and robustness are established under appropriate excitation conditions. Simulations demonstrate the performance of the developed observers and filters under noisy and noise-free measurements.

📄 PDF Abstract BibTeX arXiv:2011.02057

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)State Estimation

Similar Papers 제목 키워드 기반

Nonuniqueness and Convergence to Equivalent Solutions in Observer-based Inverse Reinforcement Learning

2022-10-28 · Jared Town, Zachary Morrison, Rushikesh Kamalapurkar

A key challenge in solving the deterministic inverse reinforcement learning (IRL) problem online and in real-time is the existence of multiple solutions. Nonuniqueness necessitates the study of the notion of equivalent s…

reinforcement-learningReinforcement Learning (RL)

KKL Observer Synthesis for Nonlinear Systems via Physics-Informed Learning

2025-01-20 · M. Umar B. Niazi, John Cao, Matthieu Barreau, Karl Henrik Johansson

This paper proposes a novel learning approach for designing Kazantzis-Kravaris/Luenberger (KKL) observers for autonomous nonlinear systems. The design of a KKL observer involves finding an injective map that transforms t…

Fault DetectionState Estimation

Learning-based Design of Luenberger Observers for Autonomous Nonlinear Systems

2022-10-04 · Muhammad Umar B. Niazi, John Cao, Xudong Sun, Amritam Das 외

Designing Luenberger observers for nonlinear systems involves the challenging task of transforming the state to an alternate coordinate system, possibly of higher dimensions, where the system is asymptotically stable and…

Pilot Performance modeling via observer-based inverse reinforcement learning

2023-07-24 · Jared Town, Zachary Morrison, Rushikesh Kamalapurkar

The focus of this paper is behavior modeling for pilots of unmanned aerial vehicles. The pilot is assumed to make decisions that optimize an unknown cost functional, which is estimated from observed trajectories using a …

reinforcement-learningReinforcement Learning

Trajectory Modeling via Random Utility Inverse Reinforcement Learning

2021-05-25 · Anselmo R. Pitombeira-Neto, Helano P. Santos, Ticiana L. Coelho da Silva, José Antonio F. de Macedo

We consider the problem of modeling trajectories of drivers in a road network from the perspective of inverse reinforcement learning. Cars are detected by sensors placed on sparsely distributed points on the street netwo…

Bayesian InferenceEconometricsreinforcement-learningReinforcement Learning+2