paper-with-me

홈 › Papers

Adversarial Imitation Learning from Video using a State Observer

2022-02-01 · Haresh Karnan, Garrett Warnell, Faraz Torabi, Peter Stone

The imitation learning research community has recently made significant progress towards the goal of enabling artificial agents to imitate behaviors from video demonstrations alone. However, current state-of-the-art approaches developed for this problem exhibit high sample complexity due, in part, to the high-dimensional nature of video observations. Towards addressing this issue, we introduce here a new algorithm called Visual Generative Adversarial Imitation from Observation using a State Observer VGAIfO-SO. At its core, VGAIfO-SO seeks to address sample inefficiency using a novel, self-supervised state observer, which provides estimates of lower-dimensional proprioceptive state representations from high-dimensional images. We show experimentally in several continuous control environments that VGAIfO-SO is more sample efficient than other IfO algorithms at learning from video-only demonstrations and can sometimes even achieve performance close to the Generative Adversarial Imitation from Observation (GAIfO) algorithm that has privileged access to the demonstrator's proprioceptive state information.

📄 PDF Abstract BibTeX arXiv:2202.00243

Code (0)

등록된 구현이 없습니다.

Tasks

continuous-controlContinuous ControlImitation Learning

Similar Papers 제목 키워드 기반

Observer-Actor: Active Vision Imitation Learning with Sparse-View Gaussian Splatting

2025-11-22 · Yilong Wang, Cheng Qian, Ruomeng Fan, Edward Johns arxiv

We propose Observer Actor (ObAct), a novel framework for active vision imitation learning in which the observer moves to optimal visual observations for the actor. We study ObAct on a dual-arm robotic system equipped wit…

Preventing Imitation Learning with Adversarial Policy Ensembles

2020-01-31 · Albert Zhan, Stas Tiomkin, Pieter Abbeel

Imitation learning can reproduce policies by observing experts, which poses a problem regarding policy privacy. Policies, such as human, or policies on deployed robots, can all be cloned without consent from the owners. …

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models

2026-03-07 · Zicheng Duan, Jiatong Xia, Zeyu Zhang, Wenbo Zhang 외 arxiv

Recent generative video world models aim to simulate visual environment evolution, allowing an observer to interactively explore the scene via camera control. However, they implicitly assume that the world only evolves w…

From Legible to Inscrutable Trajectories: (Il)legible Motion Planning Accounting for Multiple Observers

2026-02-09 · Ananya Yammanuru, Maria Lusardi, Nancy M. Amato, Katherine Driggs-Campbell arxiv

In cooperative environments, such as in factories or assistive scenarios, it is important for a robot to communicate its intentions to observers, who could be either other humans or robots. A legible trajectory allows an…

Motion Planning

ADMM based Distributed State Observer Design under Sparse Sensor Attacks

2022-09-13 · Vinaya Mary Prinse, Rachel Kalpana Kalaimani

This paper considers the design of a distributed state-observer for discrete-time Linear Time-Invariant (LTI) systems in the presence of sensor attacks. We assume there is a network of observer nodes, communicating with …

Adversarial Attack