paper-with-me

홈 › Papers

Motion Reasoning for Goal-Based Imitation Learning

2019-11-13 · De-An Huang, Yu-Wei Chao, Chris Paxton, Xinke Deng, Li Fei-Fei, Juan Carlos Niebles, Animesh Garg, Dieter Fox

We address goal-based imitation learning, where the aim is to output the symbolic goal from a third-person video demonstration. This enables the robot to plan for execution and reproduce the same goal in a completely different environment. The key challenge is that the goal of a video demonstration is often ambiguous at the level of semantic actions. The human demonstrators might unintentionally achieve certain subgoals in the demonstrations with their actions. Our main contribution is to propose a motion reasoning framework that combines task and motion planning to disambiguate the true intention of the demonstrator in the video demonstration. This allows us to robustly recognize the goals that cannot be disambiguated by previous action-based approaches. We evaluate our approach by collecting a dataset of 96 video demonstrations in a mockup kitchen environment. We show that our motion reasoning plays an important role in recognizing the actual goal of the demonstrator and improves the success rate by over 20%. We further show that by using the automatically inferred goal from the video demonstration, our robot is able to reproduce the same task in a real kitchen environment.

📄 PDF Abstract BibTeX arXiv:1911.05864

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningMotion PlanningTask and Motion Planning

Similar Papers 제목 키워드 기반

SCENIC: Scene-aware Semantic Navigation with Instruction-guided Control

2024-12-20 · Xiaohan Zhang, Sebastian Starke, Vladimir Guzov, Zhensong Zhang 외

Synthesizing natural human motion that adapts to complex environments while allowing creative control remains a fundamental challenge in motion synthesis. Existing models often fall short, either by assuming flat terrain…

Motion Synthesis

KEVER^2: Knowledge-Enhanced Visual Emotion Reasoning and Retrieval

2025-05-30 · Fanhang Man, Xiaoyue Chen, Huandong Wang, Baining Zhao 외

Understanding what emotions images evoke in their viewers is a foundational goal in human-centric visual computing. While recent advances in vision-language models (VLMs) have shown promise for visual emotion analysis (V…

Emotion RecognitionRetrieval

Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation

2025-09-21 · Liang Heng, Jiadong Xu, Yiwen Wang, Xiaoqi Li 외 arxiv

Relational object rearrangement (ROR) tasks (e.g., insert flower to vase) require a robot to manipulate objects with precise semantic and geometric reasoning. Existing approaches either rely on pre-collected demonstratio…

Object RearrangementPoint Clouds

COMMA-DEER: COmmon-sense Aware Multimodal Multitask Approach for Detection of Emotion and Emotional Reasoning in Conversations

2022-10-01 · COLING 2022 10 · Soumitra Ghosh, Gopendra Vikram Singh, Asif Ekbal, Pushpak Bhattacharyya

Mental health is a critical component of the United Nations’ Sustainable Development Goals (SDGs), particularly Goal 3, which aims to provide “good health and well-being”. The present mental health treatment gap is exace…

Common Sense Reasoning

Rational Inverse Reasoning: Few-Shot Imitation by Inferring Intent through Planning

2025-08-12 · Ben Zandonati, Tomás Lozano-Pérez, Leslie Pack Kaelbling arxiv

Humans can learn a new manipulation task from one or two demonstrations and then perform it in a new room, with new objects, under new constraints. Modern robot imitation learning, in contrast, typically needs hundreds t…