paper-with-me

홈 › Papers

Non-Cooperative Inverse Reinforcement Learning

2019-11-03 · NeurIPS 2019 12 · Xiangyuan Zhang, Kaiqing Zhang, Erik Miehling, Tamer Başar

Making decisions in the presence of a strategic opponent requires one to take into account the opponent's ability to actively mask its intended objective. To describe such strategic situations, we introduce the non-cooperative inverse reinforcement learning (N-CIRL) formalism. The N-CIRL formalism consists of two agents with completely misaligned objectives, where only one of the agents knows the true objective function. Formally, we model the N-CIRL formalism as a zero-sum Markov game with one-sided incomplete information. Through interacting with the more informed player, the less informed player attempts to both infer, and act according to, the true objective function. As a result of the one-sided incomplete information, the multi-stage game can be decomposed into a sequence of single-stage games expressed by a recursive formula. Solving this recursive formula yields the value of the N-CIRL game and the more informed player's equilibrium strategy. Another recursive formula, constructed by forming an auxiliary game, termed the dual game, yields the less informed player's strategy. Building upon these two recursive formulas, we develop a computationally tractable algorithm to approximately solve for the equilibrium strategies. Finally, we demonstrate the benefits of our N-CIRL formalism over the existing multi-agent IRL formalism via extensive numerical simulation in a novel cyber security setting.

📄 PDF Abstract BibTeX arXiv:1911.04220

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Learning Reward Models for Cooperative Trajectory Planning with Inverse Reinforcement Learning and Monte Carlo Tree Search

2022-02-14 · Karl Kurzer, Matthias Bitzer, J. Marius Zöllner

Cooperative trajectory planning methods for automated vehicles can solve traffic scenarios that require a high degree of cooperation between traffic participants. However, for cooperative systems to integrate into human-…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Individual-Level Inverse Reinforcement Learning for Mean Field Games

2022-02-13 · Yang Chen, Libo Zhang, Jiamou Liu, Shuyue Hu

The recent mean field game (MFG) formalism has enabled the application of inverse reinforcement learning (IRL) methods in large-scale multi-agent systems, with the goal of inferring reward signals that can explain demons…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Cooperative Inverse Reinforcement Learning

2016-06-09 · NeurIPS 2016 12 · Dylan Hadfield-Menell, Anca Dragan, Pieter Abbeel, Stuart Russell

For an autonomous system to be helpful to humans and to pose no unwarranted risks, it needs to align its values with those of the humans in its environment in such a way that its actions contribute to the maximization of…

Active Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Data-Driven Inverse Reinforcement Learning for Expert-Learner Zero-Sum Games

2023-01-05 · Wenqian Xue, Bosen Lian, Jialu Fan, Tianyou Chai 외

In this paper, we formulate inverse reinforcement learning (IRL) as an expert-learner interaction whereby the optimal performance intent of an expert or target agent is unknown to a learner agent. The learner observes th…

reinforcement-learningReinforcement Learning (RL)

An Efficient, Generalized Bellman Update For Cooperative Inverse Reinforcement Learning

2018-06-11 · ICML 2018 7 · Dhruv Malik, Malayandi Palaniappan, Jaime F. Fisac, Dylan Hadfield-Menell 외

Our goal is for AI systems to correctly identify and act according to their human user's objectives. Cooperative Inverse Reinforcement Learning (CIRL) formalizes this value alignment problem as a two-player game between …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)