paper-with-me

Papers

Learning Generalizable Manipulation Policies with Object-Centric 3D Representations

2023-10-22 · Yifeng Zhu, Zhenyu Jiang, Peter Stone, Yuke Zhu

We introduce GROOT, an imitation learning method for learning robust policies with object-centric and 3D priors. GROOT builds policies that generalize beyond their initial training conditions for vision-based manipulation. It constructs object-centric 3D representations that are robust toward background changes and camera views and reason over these representations using a transformer-based policy. Furthermore, we introduce a segmentation correspondence model that allows policies to generalize to new objects at test time. Through comprehensive experiments, we validate the robustness of GROOT policies against perceptual variations in simulated and real-world environments. GROOT's performance excels in generalization over background changes, camera viewpoint shifts, and the presence of new object instances, whereas both state-of-the-art end-to-end learning methods and object proposal-based approaches fall short. We also extensively evaluate GROOT policies on real robots, where we demonstrate the efficacy under very wild changes in setup. More videos and model details can be found in the appendix and the project website: https://ut-austin-rpl.github.io/GROOT .

📄 PDF Abstract BibTeX arXiv:2310.14386

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningObject

Similar Papers 제목 키워드 기반

Beyond Point-Attached Semantics: Object-Centric Semantic Fields for Generalizable Manipulation

2026-07-03 · Zheng Sun, Lerong Zhang, Zhihao Li, Zhuo Li 외 arxiv

Generalizable robot manipulation requires stable 3D understanding of functional object parts, such as handles, tool heads, openings, and graspable regions. Raw point clouds provide geometry but lack explicit part semanti…

Robot ManipulationPoint Clouds

Object-Centric Representations Improve Policy Generalization in Robot Manipulation

2025-05-16 · Alexandre Chapin, Bruno Machado, Emmanuel Dellandrea, Liming Chen

Visual representations are central to the learning and generalization capabilities of robotic manipulation policies. While existing methods rely on global or dense features, such representations often entangle task-relev…

Optical Character Recognition (OCR)Robot Manipulation

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

2026-05-14 · Jonathan Spieler, Angel Villar-Corrales, Sven Behnke arxiv

Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object-centric world models capture scene dynamics using object-level repres…

Reinforcement Learning

Deep Object-Centric Representations for Generalizable Robot Learning

2017-08-14 · Coline Devin, Pieter Abbeel, Trevor Darrell, Sergey Levine

Robotic manipulation in complex open-world scenarios requires both reliable physical manipulation skills and effective and generalizable perception. In this paper, we propose a method where general purpose pretrained vis…

ObjectReinforcement LearningReinforcement Learning (RL)

FUNCanon: Learning Pose-Aware Action Primitives via Functional Object Canonicalization for Generalizable Robotic Manipulation

2025-09-23 · Hongli Xu, Lei Zhang, Xiaoyue Hu, Boyang Zhong 외 arxiv

General-purpose robotic skills from end-to-end demonstrations often leads to task-specific policies that fail to generalize beyond the training distribution. Therefore, we introduce FunCanon, a framework that converts lo…