paper-with-me

홈 › Papers

Neural Descriptor Fields: SE(3)-Equivariant Object Representations for Manipulation

2021-12-09 · Anthony Simeonov, Yilun Du, Andrea Tagliasacchi, Joshua B. Tenenbaum, Alberto Rodriguez, Pulkit Agrawal, Vincent Sitzmann

We present Neural Descriptor Fields (NDFs), an object representation that encodes both points and relative poses between an object and a target (such as a robot gripper or a rack used for hanging) via category-level descriptors. We employ this representation for object manipulation, where given a task demonstration, we want to repeat the same task on a new object instance from the same category. We propose to achieve this objective by searching (via optimization) for the pose whose descriptor matches that observed in the demonstration. NDFs are conveniently trained in a self-supervised fashion via a 3D auto-encoding task that does not rely on expert-labeled keypoints. Further, NDFs are SE(3)-equivariant, guaranteeing performance that generalizes across all possible 3D object translations and rotations. We demonstrate learning of manipulation tasks from few (5-10) demonstrations both in simulation and on a real robot. Our performance generalizes across both object instances and 6-DoF object poses, and significantly outperforms a recent baseline that relies on 2D descriptors. Project website: https://yilundu.github.io/ndf/.

📄 PDF Abstract BibTeX arXiv:2112.05124

Code (1)

avivne/bilinear-transduction pytorch

Tasks

Object

Similar Papers 제목 키워드 기반

Equivariant Descriptor Fields: SE(3)-Equivariant Energy-Based Models for End-to-End Visual Robotic Manipulation Learning

2022-06-16 · Hyunwoo Ryu, Hong-in Lee, Jeong-Hoon Lee, Jongeun Choi

End-to-end learning for visual robotic manipulation is known to suffer from sample inefficiency, requiring large numbers of demonstrations. The spatial roto-translation equivariance, or the SE(3)-equivariance can be expl…

Local Neural Descriptor Fields: Locally Conditioned Object Representations for Manipulation

2023-02-07 · Ethan Chun, Yilun Du, Anthony Simeonov, Tomas Lozano-Perez 외

A robot operating in a household environment will see a wide range of unique and unfamiliar objects. While a system could train on many of these, it is infeasible to predict all the objects a robot will see. In this pape…

Object

RiEMann: Near Real-Time SE(3)-Equivariant Robot Manipulation without Point Cloud Segmentation

2024-03-28 · Chongkai Gao, Zhengrong Xue, Shuying Deng, Tianhai Liang 외

We present RiEMann, an end-to-end near Real-time SE(3)-Equivariant Robot Manipulation imitation learning framework from scene point cloud input. Compared to previous methods that rely on descriptor field matching, RiEMan…

Imitation LearningObjectPoint Cloud SegmentationRobot Manipulation+1

D$^3$Fields: Dynamic 3D Descriptor Fields for Zero-Shot Generalizable Rearrangement

2023-09-28 · YiXuan Wang, Mingtong Zhang, Zhuoran Li, Tarik Kelestemur 외

Scene representation is a crucial design choice in robotic manipulation systems. An ideal representation is expected to be 3D, dynamic, and semantic to meet the demands of diverse manipulation tasks. However, previous wo…

USEEK: Unsupervised SE(3)-Equivariant 3D Keypoints for Generalizable Manipulation

2022-09-28 · Zhengrong Xue, Zhecheng Yuan, Jiashun Wang, Xueqian Wang 외

Can a robot manipulate intra-category unseen objects in arbitrary poses with the help of a mere demonstration of grasping pose on a single object instance? In this paper, we try to address this intriguing challenge by us…

Keypoint DetectionObject