CaTGrasp: Learning Category-Level Task-Relevant Grasping in Clutter from Simulation
Task-relevant grasping is critical for industrial assembly, where downstream manipulation tasks constrain the set of valid grasps. Learning how to perform this task, however, is challenging, since task-relevant grasp labels are hard to define and annotate. There is also yet no consensus on proper representations for modeling or off-the-shelf tools for performing task-relevant grasps. This work proposes a framework to learn task-relevant grasping for industrial objects without the need of time-consuming real-world data collection or manual annotation. To achieve this, the entire framework is trained solely in simulation, including supervised training with synthetic label generation and self-supervised, hand-object interaction. In the context of this framework, this paper proposes a novel, object-centric canonical representation at the category level, which allows establishing dense correspondence across object instances and transferring task-relevant grasps to novel instances. Extensive experiments on task-relevant grasping of densely-cluttered industrial objects are conducted in both simulation and real-world setups, demonstrating the effectiveness of the proposed framework. Code and data are available at https://sites.google.com/view/catgrasp.
Code (1)
Tasks
Domain AdaptationDomain GeneralizationGrasp Contact PredictionGrasp GenerationIndustrial RobotsObjectPhysical SimulationsRobotic GraspingRobot Task PlanningvalidSimilar Papers 제목 키워드 기반
USEEK: Unsupervised SE(3)-Equivariant 3D Keypoints for Generalizable Manipulation
Can a robot manipulate intra-category unseen objects in arbitrary poses with the help of a mere demonstration of grasping pose on a single object instance? In this paper, we try to address this intriguing challenge by us…
Keypoint DetectionObjectTransferable Active Grasping and Real Embodied Dataset
Grasping in cluttered scenes is challenging for robot vision systems, as detection accuracy can be hindered by partial occlusion of objects. We adopt a reinforcement learning (RL) framework and 3D vision architectures to…
Reinforcement LearningReinforcement Learning (RL)Learning 6-DoF Object Poses to Grasp Category-level Objects by Language Instructions
This paper studies the task of any objects grasping from the known categories by free-form language instructions. This task demands the technique in computer vision, natural language processing, and robotics. We bring th…
ObjectObject LocalizationRobotic GraspingCategory-Level and Open-Set Object Pose Estimation for Robotics
Object pose estimation enables a variety of tasks in computer vision and robotics, including scene understanding and robotic grasping. The complexity of a pose estimation task depends on the unknown variables related to …
6D Pose Estimation6D Pose Estimation using RGBObjectPose Estimation+2HANDAL: A Dataset of Real-World Manipulable Object Categories with Pose Annotations, Affordances, and Reconstructions
We present the HANDAL dataset for category-level object pose estimation and affordance prediction. Unlike previous datasets, ours is focused on robotics-ready manipulable objects that are of the proper size and shape for…
Pose Estimation