paper-with-me

홈 › Papers

Semantically Controllable Augmentations for Generalizable Robot Learning

2024-09-02 · Zoey Chen, Zhao Mandi, Homanga Bharadhwaj, Mohit Sharma, Shuran Song, Abhishek Gupta, Vikash Kumar

Generalization to unseen real-world scenarios for robot manipulation requires exposure to diverse datasets during training. However, collecting large real-world datasets is intractable due to high operational costs. For robot learning to generalize despite these challenges, it is essential to leverage sources of data or priors beyond the robot's direct experience. In this work, we posit that image-text generative models, which are pre-trained on large corpora of web-scraped data, can serve as such a data source. These generative models encompass a broad range of real-world scenarios beyond a robot's direct experience and can synthesize novel synthetic experiences that expose robotic agents to additional world priors aiding real-world generalization at no extra cost. In particular, our approach leverages pre-trained generative models as an effective tool for data augmentation. We propose a generative augmentation framework for semantically controllable augmentations and rapidly multiplying robot datasets while inducing rich variations that enable real-world generalization. Based on diverse augmentations of robot data, we show how scalable robot manipulation policies can be trained and deployed both in simulation and in unseen real-world environments such as kitchens and table-tops. By demonstrating the effectiveness of image-text generative models in diverse real-world robotic applications, our generative augmentation framework provides a scalable and efficient path for boosting generalization in robot learning at no extra human cost.

📄 PDF Abstract BibTeX arXiv:2409.00951

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationRobot Manipulation

Similar Papers 제목 키워드 기반

Semantic-Based Explainable AI: Leveraging Semantic Scene Graphs and Pairwise Ranking to Explain Robot Failures

2021-08-08 · Devleena Das, Sonia Chernova

When interacting in unstructured human environments, occasional robot failures are inevitable. When such failures occur, everyday people, rather than trained technicians, will be the first to respond. Existing natural la…

Descriptive

MAST: Masked Augmentation Subspace Training for Generalizable Self-Supervised Priors

2023-03-07 · Chen Huang, Hanlin Goh, Jiatao Gu, Josh Susskind

Recent Self-Supervised Learning (SSL) methods are able to learn feature representations that are invariant to different data augmentations, which can then be transferred to downstream tasks of interest. However, differen…

Instance SegmentationSelf-Supervised LearningSemantic Segmentation

EgoPhys: Learning Generalizable Physics Models of Deformable Objects from Egocentric Video

2026-06-15 · Hyunjin Kim, Ri-Zhao Qiu, Guangqi Jiang, Xiaolong Wang arxiv

Humans naturally understand object physics through everyday interactions, but faithfully predicting complex deformable dynamics, such as elastic materials and fabrics, remains a major challenge for computer vision and ro…

Zero-shot Generalization

Self-supervised Visualisation of Medical Image Datasets

2024-02-22 · Ifeoma Veronica Nwabufo, Jan Niklas Böhm, Philipp Berens, Dmitry Kobak

Self-supervised learning methods based on data augmentations, such as SimCLR, BYOL, or DINO, allow obtaining semantically meaningful representations of image datasets and are widely used prior to supervised fine-tuning. …

Contrastive LearningSelf-Supervised Learning

Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation

2024-01-17 · Yinuo Zhao, Kun Wu, Tianjiao Yi, Zhiyuan Xu 외

Improving generalization is one key challenge in embodied AI, where obtaining large-scale datasets across diverse scenarios is costly. Traditional weak augmentations, such as cropping and flipping, are insufficient for i…

Data AugmentationReinforcement Learning (RL)Robot Manipulation