paper-with-me

홈 › Papers

Imitation Learning from Pixel Observations for Continuous Control

2021-09-29 · samuel cohen, Brandon Amos, Marc Peter Deisenroth, Mikael Henaff, Eugene Vinitsky, Denis Yarats

We study imitation learning using only visual observations for controlling dynamical systems with continuous states and actions. This setting is attractive due to the large amount of video data available from which agents could learn from. However, it is challenging due to $i)$ not observing the actions and $ii)$ the high-dimensional visual space. In this setting, we explore recipes for imitation learning based on adversarial learning and optimal transport. A key feature of our methods is to use representations from the RL encoder to compute imitation rewards. These recipes enable us to scale these methods to attain expert-level performance on visual continuous control tasks in the DeepMind control suite. We investigate the tradeoffs of these approaches and present a comprehensive evaluation of the key design choices. To encourage reproducible research in this area, we provide an easy-to-use implementation for benchmarking visual imitation learning, including our methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Benchmarkingcontinuous-controlContinuous ControlImitation Learning

Similar Papers 제목 키워드 기반

PixelBrax: Learning Continuous Control from Pixels End-to-End on the GPU

2025-01-16 · Trevor McInroe, Samuel Garcin

We present PixelBrax, a set of continuous control tasks with pixel observations. We combine the Brax physics engine with a pure JAX renderer, allowing reinforcement learning (RL) experiments to run end-to-end on the GPU.…

Benchmarkingcontinuous-controlContinuous ControlCPU+2

Adversarial Imitation Learning from Visual Observations using Latent Information

2023-09-29 · Vittorio Giammarino, James Queeney, Ioannis Ch. Paschalidis

We focus on the problem of imitation learning from visual observations, where the learning agent has access to videos of experts as its sole learning source. The challenges of this framework include the absence of expert…

Imitation Learning

Data-Efficient Learning of Feedback Policies from Image Pixels using Deep Dynamical Models

2015-10-08 · John-Alexander M. Assael, Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth

Data-efficient reinforcement learning (RL) in continuous state-action spaces using very high-dimensional observations remains a key challenge in developing fully autonomous systems. We consider a particularly important i…

Model-based Reinforcement LearningModel Predictive Controlreinforcement-learningReinforcement Learning+1

Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

2021-07-20 · ICLR 2022 4 · Denis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel Pinto

We present DrQ-v2, a model-free reinforcement learning (RL) algorithm for visual continuous control. DrQ-v2 builds on DrQ, an off-policy actor-critic approach that uses data augmentation to learn directly from pixels. We…

continuous-controlContinuous ControlData AugmentationGPU+4

From Pixels to Torques: Policy Learning with Deep Dynamical Models

2015-02-08 · Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth

Data-efficient learning in continuous state-action spaces using very high-dimensional observations remains a key challenge in developing fully autonomous systems. In this paper, we consider one instance of this challenge…

Model-based Reinforcement LearningModel Predictive Controlreinforcement-learningReinforcement Learning+1