paper-with-me

홈 › Papers

Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization

2016-03-01 · Chelsea Finn, Sergey Levine, Pieter Abbeel

Reinforcement learning can acquire complex behaviors from high-level specifications. However, defining a cost function that can be optimized effectively and encodes the correct task is challenging in practice. We explore how inverse optimal control (IOC) can be used to learn behaviors from demonstrations, with applications to torque control of high-dimensional robotic systems. Our method addresses two key challenges in inverse optimal control: first, the need for informative features and effective regularization to impose structure on the cost, and second, the difficulty of learning the cost function under unknown dynamics for high-dimensional continuous systems. To address the former challenge, we present an algorithm capable of learning arbitrary nonlinear cost functions, such as neural networks, without meticulous feature engineering. To address the latter challenge, we formulate an efficient sample-based approximation for MaxEnt IOC. We evaluate our method on a series of simulated tasks and real-world robotic manipulation problems, demonstrating substantial improvement over prior methods both in terms of task complexity and sample efficiency.

📄 PDF Abstract BibTeX arXiv:1603.00448

Code (4)

Alina-Samokhina/guided_cost_RL pytorch
ninawie/guided-cost-learning pytorch
nishantkr18/guided-cost-learning pytorch
opendilab/DI-engine pytorch

Tasks

Feature EngineeringReinforcement Learning

Similar Papers 제목 키워드 기반

Data-Driven Inverse Reinforcement Learning for Expert-Learner Zero-Sum Games

2023-01-05 · Wenqian Xue, Bosen Lian, Jialu Fan, Tianyou Chai 외

In this paper, we formulate inverse reinforcement learning (IRL) as an expert-learner interaction whereby the optimal performance intent of an expert or target agent is unknown to a learner agent. The learner observes th…

reinforcement-learningReinforcement Learning (RL)

Data-Driven Inverse Optimal Control for Continuous-Time Nonlinear Systems

2025-03-12 · Hamed Jabbari Asl, Eiji Uchibe

This paper introduces a novel model-free and a partially model-free algorithm for inverse optimal control (IOC), also known as inverse reinforcement learning (IRL), aimed at estimating the cost function of continuous-tim…

eXplainable AI for data driven control: an inverse optimal control approach

2025-04-15 · Federico Porcari, Donatello Materassi, Simone Formentin

Understanding the behavior of black-box data-driven controllers is a key challenge in modern control design. In this work, we propose an eXplainable AI (XAI) methodology based on Inverse Optimal Control (IOC) to obtain l…

Decision Making

Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation

2026-02-16 · Olivier Jeunen, Shashank Gupta arxiv

Off-policy evaluation (OPE) is essential for assessing ranking and recommendation systems without costly online interventions. Self-Normalised Inverse Propensity Scoring (SNIPS) is a standard tool for variance reduction …

Recommendation Systems

Optimal Control-Based Baseline for Guided Exploration in Policy Gradient Methods

2020-11-04 · Xubo Lyu, Site Li, Seth Siriya, Ye Pu 외

In this paper, a novel optimal control-based baseline function is presented for the policy gradient method in deep reinforcement learning (RL). The baseline is obtained by computing the value function of an optimal contr…

Deep Reinforcement LearningPolicy Gradient MethodsReinforcement Learning (RL)