paper-with-me

홈 › Papers

DAP: Diffusion-based Affordance Prediction for Multi-modality Storage

2024-08-31 · Haonan Chang, Kowndinya Boyalakuntla, YuHan Liu, Xinyu Zhang, Liam Schramm, Abdeslam Boularias

Solving storage problem: where objects must be accurately placed into containers with precise orientations and positions, presents a distinct challenge that extends beyond traditional rearrangement tasks. These challenges are primarily due to the need for fine-grained 6D manipulation and the inherent multi-modality of solution spaces, where multiple viable goal configurations exist for the same storage container. We present a novel Diffusion-based Affordance Prediction (DAP) pipeline for the multi-modal object storage problem. DAP leverages a two-step approach, initially identifying a placeable region on the container and then precisely computing the relative pose between the object and that region. Existing methods either struggle with multi-modality issues or computation-intensive training. Our experiments demonstrate DAP's superior performance and training efficiency over the current state-of-the-art RPDiff, achieving remarkable results on the RPDiff benchmark. Additionally, our experiments showcase DAP's data efficiency in real-world applications, an advancement over existing simulation-driven approaches. Our contribution fills a gap in robotic manipulation research by offering a solution that is both computationally efficient and capable of handling real-world variability. Code and supplementary material can be found at: https://github.com/changhaonan/DPS.git.

📄 PDF Abstract BibTeX arXiv:2409.00499

Code (1)

changhaonan/dps 공식 구현 pytorch

Tasks

Prediction

Similar Papers 제목 키워드 기반

Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning

2025-08-03 · Hanqing Wang, Zhenhao Zhang, Kaiyang Ji, Mingyu Liu 외 arxiv

3D affordance grounding aims to understand how diverse objects can be manipulated, making it a cornerstone of embodied interaction. However, prior works struggle to generalize to out-of-distribution, open-world scenarios…

ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation

2025-11-25 · Yuhan Wu, Tiantian Wei, Shuo Wang, ZhiChao Wang 외 arxiv

Interactive articulated manipulation requires long-horizon, multi-step interactions with appliances while maintaining physical consistency. Existing vision-language and diffusion-based policies struggle to generalize acr…

AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation

2025-06-24 · Ziyan Zhao, Ke Fan, He-Yang Xu, Ning Qiao 외

We present AnchorDP3, a diffusion policy framework for dual-arm robotic manipulation that achieves state-of-the-art performance in highly randomized environments. AnchorDP3 integrates three key innovations: (1) Simulator…

Multi-Task LearningSemantic SegmentationTrajectory Prediction

AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis

2026-03-09 · Xiaofei Wu, Yi Zhang, Yumeng Liu, Yuexin Ma 외 arxiv

Generating human grasping poses that accurately reflect both object geometry and user-specified interaction semantics is essential for natural hand-object interactions in AR/VR and embodied AI. However, existing semantic…

FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models

2026-01-13 · Yifan Han, Yichuan Peng, Pengfei Yi, Junyan Li 외 arxiv

Dexterous grasp synthesis must jointly satisfy functional intent and physical feasibility, yet existing pipelines often decouple semantic grounding from refinement, yielding unstable or non-functional contacts under obje…