paper-with-me

홈 › Papers

Goal-Conditioned End-to-End Visuomotor Control for Versatile Skill Primitives

2020-03-19 · Oliver Groth, Chia-Man Hung, Andrea Vedaldi, Ingmar Posner

Visuomotor control (VMC) is an effective means of achieving basic manipulation tasks such as pushing or pick-and-place from raw images. Conditioning VMC on desired goal states is a promising way of achieving versatile skill primitives. However, common conditioning schemes either rely on task-specific fine tuning - e.g. using one-shot imitation learning (IL) - or on sampling approaches using a forward model of scene dynamics i.e. model-predictive control (MPC), leaving deployability and planning horizon severely limited. In this paper we propose a conditioning scheme which avoids these pitfalls by learning the controller and its conditioning in an end-to-end manner. Our model predicts complex action sequences based directly on a dynamic image representation of the robot motion and the distance to a given target observation. In contrast to related works, this enables our approach to efficiently perform complex manipulation tasks from raw image observations without predefined control primitives or test time demonstrations. We report significant improvements in task success over representative MPC and IL baselines. We also demonstrate our model's generalisation capabilities in challenging, unseen tasks featuring visual noise, cluttered scenes and unseen object geometries.

📄 PDF Abstract BibTeX arXiv:2003.08854

Code (1)

ogroth/geeco 공식 구현 tf

Tasks

Imitation LearningMeta-LearningModel Predictive Control

Similar Papers 제목 키워드 기반

Learning Deep Parameterized Skills from Demonstration for Re-targetable Visuomotor Control

2019-10-23 · Jonathan Chang, Nishanth Kumar, Sean Hastings, Aaron Gokaslan 외

Robots need to learn skills that can not only generalize across similar problems but also be directed to a specific goal. Previous methods either train a new skill for every different goal or do not infer the specific ta…

VOFA: Visual Object Goal Pushing with Force-Adaptive Control for Humanoids

2026-05-02 · Zichao Hu, Zifan Xu, Dongsik Chang, He Yin 외 arxiv

The ability to push large objects in a goal-directed manner using onboard egocentric perception is an essential skill for humanoid robots to perform complex tasks such as material handling in warehouses. To robustly mani…

GHOST: Hierarchical Sub-Goal Policies for Generalizing Robot Manipulation

2026-06-08 · Sriram Krishna, Ben Eisner, Haotian Zhan, Ying Yuan 외 arxiv

We present GHOST, a framework for learning visuomotor manipulation policies that generalize beyond the training distribution. GHOST factorizes control into (i) a high-level policy that predicts the next sub-goal as a dis…

Robot Manipulation

LUMOS: Language-Conditioned Imitation Learning with World Models

2025-03-13 · Iman Nematollahi, Branton DeMoss, Akshay L Chandra, Nick Hawes 외

We introduce LUMOS, a language-conditioned multi-task imitation learning framework for robotics. LUMOS learns skills by practicing them over many long-horizon rollouts in the latent space of a learned world model and tra…

Imitation Learning

Learning Versatile Skills with Curriculum Masking

2024-10-23 · Yao Tang, Zhihui Xie, Zichuan Lin, Deheng Ye 외

Masked prediction has emerged as a promising pretraining paradigm in offline reinforcement learning (RL) due to its versatile masking schemes, enabling flexible inference across various downstream tasks with a unified mo…

Decision MakingOffline RLReinforcement Learning (RL)Sequential Decision Making