paper-with-me

Papers

Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation

2025-05-21 · Yihang Li, Tianle Zhang, Xuelong Wei, Jiayi Li, Lin Zhao, Dongchi Huang, Zhirui Fang, Minhua Zheng, Wenjun Dai, Xiaodong He

Robot manipulation learning from human demonstrations offers a rapid means to acquire skills but often lacks generalization across diverse scenes and object placements. This limitation hinders real-world applications, particularly in complex tasks requiring dexterous manipulation. Vision-Language-Action (VLA) paradigm leverages large-scale data to enhance generalization. However, due to data scarcity, VLA's performance remains limited. In this work, we introduce Object-Focus Actor (OFA), a novel, data-efficient approach for generalized dexterous manipulation. OFA exploits the consistent end trajectories observed in dexterous manipulation tasks, allowing for efficient policy training. Our method employs a hierarchical pipeline: object perception and pose estimation, pre-manipulation pose arrival and OFA policy execution. This process ensures that the manipulation is focused and efficient, even in varied backgrounds and positional layout. Comprehensive real-world experiments across seven tasks demonstrate that OFA significantly outperforms baseline methods in both positional and background generalization tests. Notably, OFA achieves robust performance with only 10 demonstrations, highlighting its data efficiency.

📄 PDF Abstract BibTeX arXiv:2505.15098

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectPose EstimationRobot ManipulationVision-Language-Action

Methods 이 논문이 사용한 방법론

OFA In this work, we pursue a unified paradigm for multimodal pretraining to break the scaffolds of complex task/modality-specific customization. We propose OFA, a Task-Agnostic and…

Similar Papers 제목 키워드 기반

Factored World Models for Zero-Shot Generalization in Robotic Manipulation

2022-02-10 · Ondrej Biza, Thomas Kipf, David Klee, Robert Platt 외

World models for environments with many objects face a combinatorial explosion of states: as the number of objects increases, the number of possible arrangements grows exponentially. In this paper, we learn to generalize…

Heuristic SearchObjectZero-shot Generalization

GraspFactory: A Large Object-Centric Grasping Dataset

2025-09-24 · Srinidhi Kalgundi Srinivas, Yash Shukla, Adam Arnold, Sachin Chitta arxiv

Robotic grasping is a crucial task in industrial automation, where robots are increasingly expected to handle a wide range of objects. However, a significant challenge arises when robot grasping models trained on limited…

Robotic Grasping

Towards Generalizable Robotic Data Flywheel: High-Dimensional Factorization and Composition

2026-03-26 · Yuyang Xiao, Yifei Zhou, Haoran Wang, Wenxuan Ou 외 arxiv

The lack of sufficiently diverse data, coupled with limited data efficiency, remains a major bottleneck for generalist robotic models, yet systematic strategies for collecting and curating such data are not fully explore…

Scale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic Manipulation

2026-07-23 · Yu Qi, Zhang Ye, Xinyi Xu, Yuxuan Lu 외 arxiv

Compositional generalization is essential for robot to follow diverse instructions. However, pretrained policies are known to take shortcuts, deferring to salient cues rather than grounding language. We introduce a diagn…

Decomposing the Generalization Gap in Imitation Learning for Visual Robotic Manipulation

2023-07-07 · Annie Xie, Lisa Lee, Ted Xiao, Chelsea Finn

What makes generalization hard for imitation learning in visual robotic manipulation? This question is difficult to approach at face value, but the environment from the perspective of a robot can often be decomposed into…

Imitation Learning