paper-with-me

Papers

kPAM-SC: Generalizable Manipulation Planning using KeyPoint Affordance and Shape Completion

2019-09-16 · Wei Gao, Russ Tedrake

Manipulation planning is the task of computing robot trajectories that move a set of objects to their target configuration while satisfying physically feasibility. In contrast to existing works that assume known object templates, we are interested in manipulation planning for a category of objects with potentially unknown instances and large intra-category shape variation. To achieve it, we need an object representation with which the manipulation planner can reason about both the physical feasibility and desired object configuration, while being generalizable to novel instances. The widely-used pose representation is not suitable, as representing an object with a parameterized transformation from a fixed template cannot capture large intra-category shape variation. Hence, we propose a new hybrid object representation consisting of semantic keypoint and dense geometry (a point cloud or mesh) as the interface between the perception module and motion planner. Leveraging advances in learning-based keypoint detection and shape completion, both dense geometry and keypoints can be perceived from raw sensor input. Using the proposed hybrid object representation, we formulate the manipulation task as a motion planning problem which encodes both the object target configuration and physical feasibility for a category of objects. In this way, many existing manipulation planners can be generalized to categories of objects, and the resulting perception-to-action manipulation pipeline is robust to large intra-category shape variation. Extensive hardware experiments demonstrate our pipeline can produce robot trajectories that accomplish tasks with never-before-seen objects.

📄 PDF Abstract BibTeX arXiv:1909.06980

Code (0)

등록된 구현이 없습니다.

Tasks

Keypoint DetectionMotion PlanningObject

Similar Papers 제목 키워드 기반

AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation

2025-10-01 · Anukriti Singh, Kasra Torshizi, Khuzema Habib, Kelin Yu 외 arxiv

Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background features. Existing keypoint-based approaches can focus on manipulation-cen…

3D Affordance Keypoint Detection for Robotic Manipulation

2025-11-27 · Zhiyang Liu, Ruiteng Zhao, Lei Zhou, Chengran Yuan 외 arxiv

This paper presents a novel approach for affordance-informed robotic manipulation by introducing 3D keypoints to enhance the understanding of object parts' functionality. The proposed approach provides direct information…

Semantic SegmentationAffordance DetectionKeypoint Detection

Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation

2025-03-05 · Xiaomeng Zhu, Yuyang Li, Leiyao Cui, Pengfei Li 외

Object affordance reasoning, the ability to infer object functionalities based on physical properties, is fundamental for task-oriented planning and activities in both humans and Artificial Intelligence (AI). This capabi…

ObjectObject Recognition

GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning

2026-02-04 · Guoqing Ma, Siheng Wang, Zeyu Zhang, Shan Yu 외 arxiv

Large foundation models have shown strong open-world generalization to complex problems in vision and language, but similar levels of generalization have yet to be achieved in robotics. One fundamental challenge is that …

Trajectory Planning

DexKnot: Generalizable Visuomotor Policy Learning for Dexterous Bag-Knotting Manipulation

2026-03-07 · Jiayuan Zhang, Ruihai Wu, Haojun Chen, Yuran Wang 외 arxiv

Knotting plastic bags is a common task in daily life, yet it is challenging for robots due to the bags' infinite degrees of freedom and complex physical dynamics. Existing methods often struggle in generalization to unse…