paper-with-me

홈 › Papers

AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation

2026-04-13 · Mingyang Li, Haofan Xu, Haowen Sun, Xinzhe Chen, Sihua Ren, Liqi Huang, Xinyang Sui, Chenyang Miao, Jiawei Ye, Qiongjie Cui, Zeyang Liu, Xingyu Chen, Xuguang Lan arxiv

Many everyday robot manipulation skills are affordance-dependent, with success determined by whether the robot contacts the functional object region required by the subsequent action. Current simulation data generators obtain contacts from generic grasp estimators or per-object manual contact annotations, but generic estimators rank stable grasps without task semantics and often select contacts that are misaligned with the downstream action, while manual contact annotations must be rewritten for each new object and task. To solve these challenges, we introduce AffordSim, a scalable data generator and benchmark that integrates open-vocabulary 3D affordance prediction into simulation-based trajectory generation. Given a natural-language task description, AffordSim synthesizes a task-relevant scene, emits affordance queries, grounds them on object surfaces, samples region-conditioned grasps, and selects executable candidates with motion planning. It further randomizes object pose, texture, lighting, image noise, and cross-viewpoint backgrounds for sim-to-real transfer. We instantiate AffordSim as a 50-task benchmark across diverse manipulation skills, five robot embodiments, and 500+ rigid and articulated objects. AffordSim achieves 93% of the trajectory collection success rate of manual contact annotations on affordance-critical tasks and 89% on hard composite tasks. Vision-language-action policies trained on AffordSim data transfer zero-shot to a real Franka FR3, reaching 24% average success.

📄 PDF Abstract BibTeX arXiv:2604.11674

Code (0)

등록된 구현이 없습니다.

Tasks

Robot ManipulationMotion Planning

Similar Papers 제목 키워드 기반

Affordance as general value function: A computational model

2020-10-27 · Daniel Graves, Johannes Günther, Jun Luo

General value functions (GVFs) in the reinforcement learning (RL) literature are long-term predictive summaries of the outcomes of agents following specific policies in the environment. Affordances as perceived action po…

Autonomous DrivingmodelReinforcement Learning (RL)

3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

2021-03-30 · CVPR 2021 1 · Shengheng Deng, Xun Xu, Chaozheng Wu, Ke Chen 외

The ability to understand the ways to interact with objects from visual cues, a.k.a. visual affordance, is essential to vision-guided robotic research. This involves categorizing, segmenting and reasoning of visual affor…

Affordance DetectionBenchmarkingObject

DualAfford: Learning Collaborative Visual Affordance for Dual-gripper Manipulation

2022-07-05 · Yan Zhao, Ruihai Wu, Zhehuan Chen, Yourong Zhang 외

It is essential yet challenging for future home-assistant robots to understand and manipulate diverse 3D objects in daily human environments. Towards building scalable systems that can perform diverse manipulation tasks …

3D geometry

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping

2025-07-31 · Dongming Wu, Yanping Fu, Saike Huang, Yingfei Liu 외 arxiv

General robotic grasping systems require accurate object affordance perception in diverse open-world scenarios following human instructions. However, current studies suffer from the problem of lacking reasoning-based lar…

Robot ManipulationRobotic Grasping

Are standard Object Segmentation models sufficient for Learning Affordance Segmentation?

2021-07-05 · Hugo Caselles-Dupré, Michael Garcia-Ortiz, David Filliat

Affordances are the possibilities of actions the environment offers to the individual. Ordinary objects (hammer, knife) usually have many affordances (grasping, pounding, cutting), and detecting these allow artificial ag…

ObjectSegmentationSemantic Segmentation