paper-with-me

홈 › Papers

A Novel Task-Driven Diffusion-Based Policy with Affordance Learning for Generalizable Manipulation of Articulated Objects

2025-09-18 · Hao Zhang, Zhen Kan, Weiwei Shang, Yongduan Song arxiv

Despite recent advances in dexterous manipulations, the manipulation of articulated objects and generalization across different categories remain significant challenges. To address these issues, we introduce DART, a novel framework that enhances a diffusion-based policy with affordance learning and linear temporal logic (LTL) representations to improve the learning efficiency and generalizability of articulated dexterous manipulation. Specifically, DART leverages LTL to understand task semantics and affordance learning to identify optimal interaction points. The {diffusion-based policy} then generalizes these interactions across various categories. Additionally, we exploit an optimization method based on interaction data to refine actions, overcoming the limitations of traditional diffusion policies that typically rely on offline reinforcement learning or learning from demonstrations. Experimental results demonstrate that DART outperforms most existing methods in manipulation ability, generalization performance, transfer reasoning, and robustness. For more information, visit our project website at: https://sites.google.com/view/dart0257/.

📄 PDF Abstract BibTeX arXiv:2509.14939

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

DexKnot: Generalizable Visuomotor Policy Learning for Dexterous Bag-Knotting Manipulation

2026-03-07 · Jiayuan Zhang, Ruihai Wu, Haojun Chen, Yuran Wang 외 arxiv

Knotting plastic bags is a common task in daily life, yet it is challenging for robots due to the bags' infinite degrees of freedom and complex physical dynamics. Existing methods often struggle in generalization to unse…

FUNCanon: Learning Pose-Aware Action Primitives via Functional Object Canonicalization for Generalizable Robotic Manipulation

2025-09-23 · Hongli Xu, Lei Zhang, Xiaoyue Hu, Boyang Zhong 외 arxiv

General-purpose robotic skills from end-to-end demonstrations often leads to task-specific policies that fail to generalize beyond the training distribution. Therefore, we introduce FunCanon, a framework that converts lo…

Affordance Learning for End-to-End Visuomotor Robot Control

2019-03-10 · Aleksi Hämäläinen, Karol Arndt, Ali Ghadirzadeh, Ville Kyrki

Training end-to-end deep robot policies requires a lot of domain-, task-, and hardware-specific data, which is often costly to provide. In this work, we propose to tackle this issue by employing a deep neural network wit…

Dataset Generation

AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation

2025-06-24 · Ziyan Zhao, Ke Fan, He-Yang Xu, Ning Qiao 외

We present AnchorDP3, a diffusion policy framework for dual-arm robotic manipulation that achieves state-of-the-art performance in highly randomized environments. AnchorDP3 integrates three key innovations: (1) Simulator…

Multi-Task LearningSemantic SegmentationTrajectory Prediction

AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation

2025-10-01 · Anukriti Singh, Kasra Torshizi, Khuzema Habib, Kelin Yu 외 arxiv

Vision-based robot learning often relies on dense image or point-cloud inputs, which are computationally heavy and entangle irrelevant background features. Existing keypoint-based approaches can focus on manipulation-cen…