paper-with-me

Papers

RAIL: A modular framework for Reinforcement-learning-based Adversarial Imitation Learning

2021-05-08 · Eddy Hudson, Garrett Warnell, Peter Stone

While Adversarial Imitation Learning (AIL) algorithms have recently led to state-of-the-art results on various imitation learning benchmarks, it is unclear as to what impact various design decisions have on performance. To this end, we present here an organizing, modular framework called Reinforcement-learning-based Adversarial Imitation Learning (RAIL) that encompasses and generalizes a popular subclass of existing AIL approaches. Using the view espoused by RAIL, we create two new IfO (Imitation from Observation) algorithms, which we term SAIfO: SAC-based Adversarial Imitation from Observation and SILEM (Skeletal Feature Compensation for Imitation Learning with Embodiment Mismatch). We go into greater depth about SILEM in a separate technical report. In this paper, we focus on SAIfO, evaluating it on a suite of locomotion tasks from OpenAI Gym, and showing that it outperforms contemporaneous RAIL algorithms that perform IfO.

📄 PDF Abstract BibTeX arXiv:2105.03756

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningOpenAI Gymreinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Diffusion-Reward Adversarial Imitation Learning

2024-05-25 · Chun-Mao Lai, Hsiang-Chun Wang, Ping-Chun Hsieh, Yu-Chiang Frank Wang 외

Imitation learning aims to learn a policy from observing expert demonstrations without access to reward signals from environments. Generative adversarial imitation learning (GAIL) formulates imitation learning as adversa…

Imitation Learning

Task-Relevant Adversarial Imitation Learning

2019-10-02 · Konrad Zolna, Scott Reed, Alexander Novikov, Sergio Gomez Colmenarejo 외

We show that a critical vulnerability in adversarial imitation is the tendency of discriminator networks to learn spurious associations between visual features and expert labels. When the discriminator focuses on task-ir…

Imitation Learning

Contrail-to-Flight Attribution Using Ground Visible Cameras and Flight Surveillance Data

2025-10-19 · Ramon Dalmau, Gabriel Jarry, Philippe Very arxiv

Aviation's non-CO2 effects, particularly contrails, are a significant contributor to its climate impact. Persistent contrails can evolve into cirrus-like clouds that trap outgoing infrared radiation, with radiative forci…

Lightweight Safety Guardrails via Synthetic Data and RL-guided Adversarial Training

2025-07-11 · Aleksei Ilin, Gor Matevosyan, Xueying Ma, Vladimir Eremin 외

We introduce a lightweight yet highly effective safety guardrail framework for language models, demonstrating that small-scale language models can achieve, and even surpass, the performance of larger counterparts in cont…

Generative Adversarial NetworkSynthetic Data Generation

RAIL: Risk-Averse Imitation Learning

2017-07-20 · Anirban Santara, Abhishek Naik, Balaraman Ravindran, Dipankar Das 외

Imitation learning algorithms learn viable policies by imitating an expert's behavior when reward signals are not available. Generative Adversarial Imitation Learning (GAIL) is a state-of-the-art algorithm for learning p…

Autonomous Drivingcontinuous-controlContinuous ControlImitation Learning