paper-with-me

홈 › Papers

A Ranking Game for Imitation Learning

2022-02-07 · Harshit Sikchi, Akanksha Saran, Wonjoon Goo, Scott Niekum

We propose a new framework for imitation learning -- treating imitation as a two-player ranking-based game between a policy and a reward. In this game, the reward agent learns to satisfy pairwise performance rankings between behaviors, while the policy agent learns to maximize this reward. In imitation learning, near-optimal expert data can be difficult to obtain, and even in the limit of infinite data cannot imply a total ordering over trajectories as preferences can. On the other hand, learning from preferences alone is challenging as a large number of preferences are required to infer a high-dimensional reward function, though preference data is typically much easier to collect than expert demonstrations. The classical inverse reinforcement learning (IRL) formulation learns from expert demonstrations but provides no mechanism to incorporate learning from offline preferences and vice versa. We instantiate the proposed ranking-game framework with a novel ranking loss giving an algorithm that can simultaneously learn from expert demonstrations and preferences, gaining the advantages of both modalities. Our experiments show that the proposed method achieves state-of-the-art sample efficiency and can solve previously unsolvable tasks in the Learning from Observation (LfO) setting. Project video and code can be found at https://hari-sikchi.github.io/rank-game/

📄 PDF Abstract BibTeX arXiv:2202.03481

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learning

Similar Papers 제목 키워드 기반

On the Limitations of Elo: Real-World Games, are Transitive, not Additive

2022-06-21 · Quentin Bertrand, Wojciech Marian Czarnecki, Gauthier Gidel

Real-world competitive games, such as chess, go, or StarCraft II, rely on Elo models to measure the strength of their players. Since these games are not fully transitive, using Elo implicitly assumes they have a strong t…

StarcraftStarcraft II

Micro- and Macro-Level Churn Analysis of Large-Scale Mobile Games

2019-01-14 · Xi Liu, Muhe Xie, Xidao Wen, Rui Chen 외

As mobile devices become more and more popular, mobile gaming has emerged as a promising market with billion-dollar revenues. A variety of mobile game platforms and services have been developed around the world. A critic…

Attribute

A Bayesian Approximation Method for Online Ranking

2011-01-01 · Journal of Machine Learning Research 2011 1 · Ruby C. Weng, Chih-Jen Lin

This paper describes a Bayesian approximation method to obtain online ranking algorithms for games with multiple teams and multiple players. Recently for Internet games large online ranking systems are much needed. We …

Reasoning, Memorization, and Fine-Tuning Language Models for Non-Cooperative Games

2024-10-18 · Yunhao Yang, Leonard Berthellemy, Ufuk Topcu

We develop a method that integrates the tree of thoughts and multi-agent framework to enhance the capability of pre-trained language models in solving complex, unfamiliar games. The method decomposes game-solving into fo…

Language ModelingLanguage ModellingMemorization

On the Convergence of No-Regret Dynamics in Information Retrieval Games with Proportional Ranking Functions

2024-05-19 · Omer Madmon, Idan Pipano, Itamar Reinman, Moshe Tennenholtz

Publishers who publish their content on the web act strategically, in a behavior that can be modeled within the online learning framework. Regret, a central concept in machine learning, serves as a canonical measure for …

Information RetrievalRetrieval