paper-with-me

홈 › Papers

AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation

2025-10-01 · Anukriti Singh, Kasra Torshizi, Khuzema Habib, Kelin Yu, Ruohan Gao, Pratap Tokekar arxiv

Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background features. Existing keypoint-based approaches can focus on manipulation-centric features and be lightweight, but either depend on manual heuristics or task-coupled selection, limiting scalability and semantic understanding. To address this, we propose AFFORD2ACT, an affordance-guided framework that distills a minimal set of semantic 2D keypoints from a text prompt and a single image. AFFORD2ACT follows a three-stage pipeline: affordance filtering, category-level keypoint construction, and transformer-based policy learning with embedded gating to reason about the most relevant keypoints, yielding a compact 38-dimensional state policy that can be trained in 15 minutes, which performs well in real-time without proprioception or dense representations. Across diverse real-world manipulation tasks, AFFORD2ACT consistently improves data efficiency, achieving an 82% success rate on unseen objects, novel categories, backgrounds, and distractors.

📄 PDF Abstract BibTeX arXiv:2510.01433

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi-Keypoint Affordance Representation for Functional Dexterous Grasping

2025-02-27 · Fan Yang, Dongsheng Luo, Wenrui Chen, Jiacheng Lin 외

Functional dexterous grasping requires precise hand-object interaction, going beyond simple gripping. Existing affordance-based methods primarily predict coarse interaction regions and cannot directly constrain the grasp…

Weakly Supervised Affordance Detection

2017-07-01 · CVPR 2017 7 · Johann Sawatzky, Abhilash Srikantha, Juergen Gall

Localizing functional regions of objects or affordances is an important aspect of scene understanding and relevant for many robotics applications. In this work, we introduce a pixel-wise annotated affordance dataset of 3…

Affordance DetectionObjectScene Understanding

3D Affordance Keypoint Detection for Robotic Manipulation

2025-11-27 · Zhiyang Liu, Ruiteng Zhao, Lei Zhou, Chengran Yuan 외 arxiv

This paper presents a novel approach for affordance-informed robotic manipulation by introducing 3D keypoints to enhance the understanding of object parts' functionality. The proposed approach provides direct information…

Semantic SegmentationAffordance DetectionKeypoint Detection

Affordance-Guided Reinforcement Learning via Visual Prompting

2024-07-14 · Olivia Y. Lee, Annie Xie, Kuan Fang, Karl Pertsch 외

Robots equipped with reinforcement learning (RL) have the potential to learn a wide range of skills solely from a reward signal. However, obtaining a robust and dense reward signal for general manipulation tasks remains …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Visual Prompting+1

Affostruction: 3D Affordance Grounding with Generative Reconstruction

2026-01-14 · Chunghyun Park, Seunghyeon Lee, Minsu Cho arxiv

This paper addresses the problem of affordance grounding from RGBD images of an object, which aims to localize surface regions corresponding to a text query that describes an action on the object. While existing methods …

3D Reconstruction