paper-with-me

홈 › Papers

SimToolReal: An Object-Centric Policy for Zero-Shot Dexterous Tool Manipulation

2026-02-18 · Kushal Kedia, Tyler Ga Wei Lum, Jeannette Bohg, C. Karen Liu arxiv

The ability to manipulate tools significantly expands the set of tasks a robot can perform. Yet, tool manipulation represents a challenging class of dexterity, requiring grasping thin objects, in-hand object rotations, and forceful interactions. Since collecting teleoperation data for these behaviors is challenging, sim-to-real reinforcement learning (RL) is a promising alternative. However, prior approaches typically require substantial engineering effort to model objects and tune reward functions for each task. In this work, we propose SimToolReal, taking a step towards generalizing sim-to-real RL policies for tool manipulation. Instead of focusing on a single object and task, we procedurally generate a large variety of tool-like object primitives in simulation and train a single RL policy with the universal goal of manipulating each object to random goal poses. This approach enables SimToolReal to perform general dexterous tool manipulation at test-time without any object or task-specific training. We demonstrate that SimToolReal outperforms prior retargeting and fixed-grasp methods by 37% while matching the performance of specialist RL policies trained on specific target objects and tasks. Finally, we show that SimToolReal generalizes across a diverse set of everyday tools, achieving strong zero-shot performance over 120 real-world rollouts spanning 24 tasks, 12 object instances, and 6 tool categories.

📄 PDF Abstract BibTeX arXiv:2602.16863

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Object-Centric Residual RL for Zero-Shot Sim-to-Real VLA Enhancement

2026-06-17 · Kinam Kim, Namiko Saito, Heecheol Kim, Katsushi Ikeuchi 외 arxiv

Vision-Language-Action (VLA) models can generalize across diverse manipulation tasks, but their imitation-learning-based policies remain brittle in precise physical interactions due to compounding execution errors; Can a…

Reinforcement Learning

Zero-shot Object-Centric Instruction Following: Integrating Foundation Models with Traditional Navigation

2024-11-12 · Sonia Raychaudhuri, Duy Ta, Katrina Ashton, Angel X. Chang 외

Large scale scenes such as multifloor homes can be robustly and efficiently mapped with a 3D graph of landmarks estimated jointly with robot poses in a factor graph, a technique commonly used in commercial robots such as…

Instruction FollowingObjectVision-Language Navigation

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos

2026-05-24 · Zhi Wang, Botao He, Kelin Yu, Seungjae Lee 외 arxiv

Human egocentric video captures rich manipulation demonstrations without any robot hardware, yet transferring these skills to robots remains challenging due to the embodiment gap between human and robot in both visual ap…

ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration

2026-04-09 · Yanwen Zou, Chenyang Shi, Wenye Yu, Han Xue 외 arxiv

Large-scale real-world robot data collection is a prerequisite for bringing robots into everyday deployment. However, existing pipelines often rely on specialized handheld devices to bridge the embodiment gap, which not …

Robot Manipulation

Zero-Shot Object-Centric Representation Learning

2024-08-17 · Aniket Didolkar, Andrii Zadaianchuk, Anirudh Goyal, Mike Mozer 외

The goal of object-centric representation learning is to decompose visual scenes into a structured representation that isolates the entities. Recent successes have shown that object-centric representation learning can be…

ObjectObject DiscoveryRepresentation LearningZero-shot Generalization