paper-with-me

홈 › Papers

Actron3D: Learning Actionable Neural Functions from Videos for Transferable Robotic Manipulation

2025-10-14 · Anran Zhang, Hanzhi Chen, Yannick Burkhardt, Yao Zhong, Johannes Betz, Helen Oleynikova, Stefan Leutenegger arxiv

We present Actron3D, a framework that enables robots to acquire transferable 6-DoF manipulation skills from just a few monocular, uncalibrated, RGB-only human videos. At its core lies the Neural Affordance Function, a compact object-centric representation that distills actionable cues from diverse uncalibrated videos-geometry, visual appearance, and affordance-into a lightweight neural network, forming a memory bank of manipulation skills. During deployment, we adopt a pipeline that retrieves relevant affordance functions and transfers precise 6-DoF manipulation policies via coarse-to-fine optimization, enabled by continuous queries to the multimodal features encoded in the neural functions. Experiments in both simulation and the real world demonstrate that Actron3D significantly outperforms prior methods, achieving a 14.9 percentage point improvement in average success rate across 13 tasks while requiring only 2-3 demonstration videos per task.

📄 PDF Abstract BibTeX arXiv:2510.12971

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Predicting Chemical Reaction Outcomes Based on Electron Movements Using Machine Learning

2025-03-13 · Shuan Chen, Kye Sung Park, Taewan Kim, Sunkyu Han 외

Accurately predicting chemical reaction outcomes and potential byproducts is a fundamental task of modern chemistry, enabling the efficient design of synthetic pathways and driving progress in chemical science. Reaction …

Interactron: Embodied Adaptive Object Detection

2022-02-01 · CVPR 2022 1 · Klemen Kotar, Roozbeh Mottaghi

Over the years various methods have been proposed for the problem of object detection. Recently, we have witnessed great strides in this domain owing to the emergence of powerful deep neural networks. However, there are …

Objectobject-detectionObject Detection

VideoWorld 2: Learning Transferable Knowledge from Real-world Videos

2026-02-10 · Zhongwei Ren, Yunchao Wei, Xiao Yu, Guixun Luo 외 arxiv

Learning transferable knowledge from unlabeled video data and applying it in new environments is a fundamental capability of intelligent agents. This work presents VideoWorld 2, which extends VideoWorld and offers the fi…

Video Generation

GLOVER++: Unleashing the Potential of Affordance Learning from Human Behaviors for Robotic Manipulation

2025-05-17 · Teli Ma, Jia Zheng, Zifan Wang, Ziyao Gao 외

Learning manipulation skills from human demonstration videos offers a promising path toward generalizable and interpretable robotic intelligence-particularly through the lens of actionable affordances. However, transferr…

Benchmarking

Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

2021-04-15 · Yevgen Chebotar, Karol Hausman, Yao Lu, Ted Xiao 외

We consider the problem of learning useful robotic skills from previously collected offline data without access to manually specified rewards or additional online exploration, a setting that is becoming increasingly impo…

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)