paper-with-me

홈 › Papers

Dexterous Point Policy: Learning Point-based Dexterous Hand Policies from Human Demonstrations

2026-06-09 · Beomjun Kim, Seong Hyeon Park, Seunghoon Sim, Seungjun Moon, Sanghyeok Lee, Jinwoo Shin arxiv

Robotic foundation models pre-trained on human demonstration videos have shown promise, but a significant embodiment gap remains when the resulting policies are deployed on real robots. A common remedy is to fine-tune these models on robot-specific demonstrations. However, robot data collection can be prohibitively expensive and time-consuming, which is particularly acute in dexterous manipulation, e.g., teleoperating a multi-fingered hand for even a single atomic task can take days. To address this, we introduce Dexterous Point Policy, a framework that learns dexterous manipulation policies directly from human videos and requires no robot demonstrations. Our core insight is that a unified 3D keypoint representation can bridge human and robot embodiments when used for both observations and actions. Specifically, we extract 3D keypoints of task-relevant objects and human hands from raw videos, and train an autoregressive transformer over these keypoints. We observe that at the keypoint level, specifically the wrist and fingertips, human and robot behaviors closely align, enabling direct policy transfer. On a suite of real-robot tasks spanning pick-and-place and tool use, Dexterous Point Policy attains 75.0% success, whereas a state-of-the-art VLA baseline reaches only 1.0%. Furthermore, our method generalizes strongly to unseen scenarios, including multi-object environments and novel object categories.

📄 PDF Abstract BibTeX arXiv:2606.10614

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DexPoint: Generalizable Point Cloud Reinforcement Learning for Sim-to-Real Dexterous Manipulation

2022-11-17 · Yuzhe Qin, Binghao Huang, Zhao-Heng Yin, Hao Su 외

We propose a sim-to-real framework for dexterous manipulation which can generalize to new objects of the same category in the real world. The key of our framework is to train the manipulation policy with point cloud inpu…

reinforcement-learningReinforcement Learning (RL)

TransDex: Pre-training Visuo-Tactile Policy with Point Cloud Reconstruction for Dexterous Manipulation of Transparent Objects

2026-03-14 · Fengguan Li, Yifan Ma, Chen Qian, Wentao Rao 외 arxiv

Dexterous manipulation enables complex tasks but suffers from self-occlusion, severe depth noise, and depth information loss when manipulating transparent objects. To solve this problem, this paper proposes TransDex, a 3…

Point Clouds

MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos

2025-04-08 · Alexey Gavryushin, Xi Wang, Robert J. S. Malate, Chenyu Yang 외

Large-scale egocentric video datasets capture diverse human activities across a wide range of scenarios, offering rich and detailed insights into how humans interact with objects, especially those that require fine-grain…

CordViP: Correspondence-based Visuomotor Policy for Dexterous Manipulation in Real-World

2025-02-12 · Yankai Fu, Qiuxuan Feng, Ning Chen, Zichen Zhou 외

Achieving human-level dexterity in robots is a key objective in the field of robotic manipulation. Recent advancements in 3D-based imitation learning have shown promising results, providing an effective pathway to achiev…

6D Pose EstimationImitation LearningPose Estimation

TopoRetarget: Interaction-Preserving Retargeting for Dexterous Manipulation

2026-06-15 · Jielin Wu, Shenzhe Yao, Guanqi He, Xiaohan Liu 외 arxiv

Human hand-object demonstrations provide dense reference motions for training dexterous manipulation reinforcement learning (RL) policies through reference tracking. However, to use such demonstrations for RL policy lear…

Reinforcement Learning