paper-with-me

홈 › Papers

Pose-Agnostic Robotic Functional Grasping via Observation-Action Canonicalization

2026-06-19 · Le Qiu, Cole Harrison, Jiankai Sun, Yao Liu, Suning Huang, Qianzhong Chen, Yang You, Marco Pavone arxiv

Functional robotic grasping requires a policy that generalizes across diverse object geometries and poses while maintaining task-specific contact precision. We study this challenge through mug-handle grasping, where thin handles, instance variation, and upright or inverted placements make both perception and control sensitive to object configuration. Grasp pose detection methods operate open-loop and are sensitive to estimation errors on thin handle structures. Learned visuomotor policies must implicitly learn to handle the coupled variation in visual appearance and action direction induced by different object placements, limiting generalization. We propose AnyMug, a canonicalized visuomotor reinforcement learning framework for functional grasping that trains a single closed-loop policy entirely in simulation and deploys it zero-shot on a real robot. AnyMug introduces observation-action canonicalization, which transforms both the depth observation and the predicted end-effector action into a shared object-centric frame. The policy therefore sees a consistent mug-centered view and emits actions in a canonical direction regardless of mug placement, allowing the same grasping behavior to be reused across configurations. A handle-aware reward further encourages precise approach, gripper alignment, and opposing-finger placement, while a pose curriculum and domain randomization improve training stability and sim-to-real transfer. In simulation, AnyMug achieves over 93% success rate on both unseen upright and inverted mugs and transfers zero-shot to a real Franka Panda, reaching 80% success rate on 5 held-out physical mugs across both pose categories.

📄 PDF Abstract BibTeX arXiv:2606.21148

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningRobotic Grasping

Similar Papers 제목 키워드 기반

GAMMA: Graspability-Aware Mobile MAnipulation Policy Learning based on Online Grasping Pose Fusion

2023-09-27 · Jiazhao Zhang, Nandiraju Gireesh, Jilong Wang, Xiaomeng Fang 외

Mobile manipulation constitutes a fundamental task for robotic assistants and garners significant attention within the robotics community. A critical challenge inherent in mobile manipulation is the effective observation…

FunGrasp: Functional Grasping for Diverse Dexterous Hands

2024-11-24 · Linyi Huang, HUI ZHANG, Zijian Wu, Sammy Christen 외

Functional grasping is essential for humans to perform specific tasks, such as grasping scissors by the finger holes to cut materials or by the blade to safely hand them over. Enabling dexterous robot hands with function…

Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Grasping in Dexterous Robotics

2024-06-30 · Fan Yang, Wenrui Chen, Kailun Yang, Haoran Lin 외

To enable robots to use tools, the initial step is teaching robots to employ dexterous gestures for touching specific areas precisely where tasks are performed. Affordance features of objects serve as a bridge in the fun…

Human-Object Interaction DetectionObject

IFG: Internet-Scale Guidance for Functional Grasping Generation

2025-11-12 · Ray Muxin Liu, Mingxuan Li, Kenneth Shaw, Deepak Pathak arxiv

Large Vision Models trained on internet-scale data have demonstrated strong capabilities in segmenting and semantically understanding object parts, even in cluttered, crowded scenes. However, while these models can direc…

Point Clouds

Multi-Keypoint Affordance Representation for Functional Dexterous Grasping

2025-02-27 · Fan Yang, Dongsheng Luo, Wenrui Chen, Jiacheng Lin 외

Functional dexterous grasping requires precise hand-object interaction, going beyond simple gripping. Existing affordance-based methods primarily predict coarse interaction regions and cannot directly constrain the grasp…