paper-with-me

Papers

Planning-Guided Diffusion Policy Learning for Generalizable Contact-Rich Bimanual Manipulation

2024-12-03 · Xuanlin Li, Tong Zhao, Xinghao Zhu, Jiuguang Wang, Tao Pang, Kuan Fang

Contact-rich bimanual manipulation involves precise coordination of two arms to change object states through strategically selected contacts and motions. Due to the inherent complexity of these tasks, acquiring sufficient demonstration data and training policies that generalize to unseen scenarios remain a largely unresolved challenge. Building on recent advances in planning through contacts, we introduce Generalizable Planning-Guided Diffusion Policy Learning (GLIDE), an approach that effectively learns to solve contact-rich bimanual manipulation tasks by leveraging model-based motion planners to generate demonstration data in high-fidelity physics simulation. Through efficient planning in randomized environments, our approach generates large-scale and high-quality synthetic motion trajectories for tasks involving diverse objects and transformations. We then train a task-conditioned diffusion policy via behavior cloning using these demonstrations. To tackle the sim-to-real gap, we propose a set of essential design options in feature extraction, task representation, action prediction, and data augmentation that enable learning robust prediction of smooth action sequences and generalization to unseen scenarios. Through experiments in both simulation and the real world, we demonstrate that our approach can enable a bimanual robotic system to effectively manipulate objects of diverse geometries, dimensions, and physical properties. Website: https://glide-manip.github.io/

📄 PDF Abstract BibTeX arXiv:2412.02676

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Query-Centric Diffusion Policy for Generalizable Robotic Assembly

2025-09-23 · Ziyi Xu, Haohong Lin, Shiqi Liu, Ding Zhao arxiv

The robotic assembly task poses a key challenge in building generalist robots due to the intrinsic complexity of part interactions and the sensitivity to noise perturbations in contact-rich settings. The assembly agent i…

HeteroGenManip: Generalizable Manipulation For Heterogeneous Object Interactions

2026-05-11 · Zhenhao Shen, Zeming Yang, Yue Chen, Yuran Wang 외 arxiv

Generalizable manipulation involving cross-type object interactions is a critical yet challenging capability in robotics. To reliably accomplish such tasks, robots must address two fundamental challenges: "where to manip…

Trajectory Planning

DexHandDiff: Interaction-aware Diffusion Planning for Adaptive Dexterous Manipulation

2024-11-27 · CVPR 2025 1 · Zhixuan Liang, Yao Mu, Yixiao Wang, Tianxing Chen 외

Dexterous manipulation with contact-rich interactions is crucial for advanced robotics. While recent diffusion-based planning approaches show promise for simple manipulation tasks, they often produce unrealistic ghost st…

Contact-rich Manipulation

EquiContact: A Hierarchical SE(3) Vision-to-Force Equivariant Policy for Spatially Generalizable Contact-rich Tasks

2025-07-15 · Joohwan Seo, Arvind Kruthiventy, Soomi Lee, Megan Teng 외 arxiv

This paper presents a framework for learning vision-based robotic policies for contact-rich manipulation tasks that generalize spatially across task configurations. We focus on achieving robust spatial generalization of …

Semantic-Contact Fields for Category-Level Generalizable Tactile Tool Manipulation

2026-02-14 · Kevin Yuchen Ma, Heng Zhang, Weisi Lin, Mike Zheng Shou 외 arxiv

Generalizing tool manipulation requires both semantic planning and precise physical control. Modern generalist robot policies, such as Vision-Language-Action (VLA) models, often lack the physical grounding required for c…