paper-with-me

Papers

Query-Centric Diffusion Policy for Generalizable Robotic Assembly

2025-09-23 · Ziyi Xu, Haohong Lin, Shiqi Liu, Ding Zhao arxiv

The robotic assembly task poses a key challenge in building generalist robots due to the intrinsic complexity of part interactions and the sensitivity to noise perturbations in contact-rich settings. The assembly agent is typically designed in a hierarchical manner: high-level multi-part reasoning and low-level precise control. However, implementing such a hierarchical policy is challenging in practice due to the mismatch between high-level skill queries and low-level execution. To address this, we propose the Query-centric Diffusion Policy (QDP), a hierarchical framework that bridges high-level planning and low-level control by utilizing queries comprising objects, contact points, and skill information. QDP introduces a query-centric mechanism that identifies task-relevant components and uses them to guide low-level policies, leveraging point cloud observations to improve the policy's robustness. We conduct comprehensive experiments on the FurnitureBench in both simulation and real-world settings, demonstrating improved performance in skill precision and long-horizon success rate. In the challenging insertion and screwing tasks, QDP improves the skill-wise success rate by over 50% compared to baselines without structured queries.

📄 PDF Abstract BibTeX arXiv:2509.18686

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

3D Flow Diffusion Policy: Visuomotor Policy Learning via Generating Flow in 3D Space

2025-09-23 · Sangjun Noh, Dongwoo Nam, Kangmin Kim, Geonhyup Lee 외 arxiv

Learning robust visuomotor policies that generalize across diverse objects and interaction dynamics remains a central challenge in robotic manipulation. Most existing approaches rely on direct observation-to-action mappi…

FUNCanon: Learning Pose-Aware Action Primitives via Functional Object Canonicalization for Generalizable Robotic Manipulation

2025-09-23 · Hongli Xu, Lei Zhang, Xiaoyue Hu, Boyang Zhong 외 arxiv

General-purpose robotic skills from end-to-end demonstrations often leads to task-specific policies that fail to generalize beyond the training distribution. Therefore, we introduce FunCanon, a framework that converts lo…

Unifying Object-Centric World Models and Diffusion Policy: A Hierarchical Framework for Multi-Stage Robotic Tasks

2026-06-07 · Raktim Gautam Goswami, Prashanth Krishnamurthy, Yann LeCun, Farshad Khorrami arxiv

Visual world models have shown great potential in learning complex system dynamics. Recent advancements leverage these models as transition functions within Model Predictive Control (MPC) frameworks to solve various cont…

Generalizable Humanoid Manipulation with 3D Diffusion Policies

2024-10-14 · Yanjie Ze, Zixuan Chen, Wenhao Wang, Tianyi Chen 외

Humanoid robots capable of autonomous operation in diverse environments have long been a goal for roboticists. However, autonomous manipulation by humanoid robots has largely been restricted to one specific scene, primar…

Camera CalibrationPoint Cloud Segmentation

Unpacking the Individual Components of Diffusion Policy

2024-11-27 · Xiu Yuan

Imitation Learning presents a promising approach for learning generalizable and complex robotic skills. The recently proposed Diffusion Policy generates robot action sequences through a conditional denoising diffusion pr…

DenoisingImitation Learning