paper-with-me

Papers

Cost Function Estimation Using Inverse Reinforcement Learning with Minimal Observations

2025-05-13 · Sarmad Mehrdad, Avadesh Meduri, Ludovic Righetti

We present an iterative inverse reinforcement learning algorithm to infer optimal cost functions in continuous spaces. Based on a popular maximum entropy criteria, our approach iteratively finds a weight improvement step and proposes a method to find an appropriate step size that ensures learned cost function features remain similar to the demonstrated trajectory features. In contrast to similar approaches, our algorithm can individually tune the effectiveness of each observation for the partition function and does not need a large sample set, enabling faster learning. We generate sample trajectories by solving an optimal control problem instead of random sampling, leading to more informative trajectories. The performance of our method is compared to two state of the art algorithms to demonstrate its benefits in several simulated environments.

📄 PDF Abstract BibTeX arXiv:2505.08619

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Online Observer-Based Inverse Reinforcement Learning

2020-11-03 · Ryan Self, Kevin Coleman, He Bai, Rushikesh Kamalapurkar

In this paper, a novel approach to the output-feedback inverse reinforcement learning (IRL) problem is developed by casting the IRL problem, for linear systems with quadratic cost functions, as a state estimation problem…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)State Estimation

Toward Global Intent Inference for Human Motion by Inverse Reinforcement Learning

2026-03-08 · Sarmad Mehrdad, Maxime Sabbah, Vincent Bonnet, Ludovic Righetti arxiv

This paper investigates whether a single, unified cost function can explain and predict human reaching movements, in contrast with existing approaches that rely on subject- or posture-specific optimization criteria. Usin…

Reinforcement Learning

Learning Human Reaching Optimality Principles from Minimal Observation Inverse Reinforcement Learning

2025-09-30 · Sarmad Mehrdad, Maxime Sabbah, Vincent Bonnet, Ludovic Righetti arxiv

This paper investigates the application of Minimal Observation Inverse Reinforcement Learning (MO-IRL) to model and predict human arm-reaching movements with time-varying cost weights. Using a planar two-link biomechanic…

Reinforcement Learning

Provably Efficient Exploration in Inverse Constrained Reinforcement Learning

2024-09-24 · Bo Yue, Jian Li, Guiliang Liu

Optimizing objective functions subject to constraints is fundamental in many real-world applications. However, these constraints are often not readily defined and must be inferred from expert agent behaviors, a problem k…

Efficient Explorationreinforcement-learningReinforcement Learning

Finite-Sample Bounds for Adaptive Inverse Reinforcement Learning using Passive Langevin Dynamics

2023-04-18 · Luke Snow, Vikram Krishnamurthy

This paper provides a finite-sample analysis of a passive stochastic gradient Langevin dynamics (PSGLD) algorithm. This algorithm is designed to achieve adaptive inverse reinforcement learning (IRL). Adaptive IRL aims to…

Density Estimationreinforcement-learningReinforcement Learning