paper-with-me

홈 › Papers

OXE-AugE: A Large-Scale Robot Augmentation of OXE for Scaling Cross-Embodiment Policy Learning

2025-12-15 · Guanhua Ji, Harsha Polavaram, Lawrence Yunliang Chen, Sandeep Bajamahal, Zehan Ma, Simeon Adebola, Chenfeng Xu, Ken Goldberg arxiv

Large and diverse datasets are needed for training generalist robot policies that have potential to control a variety of robot embodiments -- robot arm and gripper combinations -- across diverse tasks and environments. As re-collecting demonstrations and retraining for each new hardware platform are prohibitively costly, we show that existing robot data can be augmented for transfer and generalization. The Open X-Embodiment (OXE) dataset, which aggregates demonstrations from over 60 robot datasets, has been widely used as the foundation for training generalist policies. However, it is highly imbalanced: the top four robot types account for over 85\% of its real data, which risks overfitting to robot-scene combinations. We present AugE-Toolkit, a scalable robot augmentation pipeline, and OXE-AugE, a high-quality open-source dataset that augments OXE with 9 different robot embodiments. OXE-AugE provides over 4.4 million trajectories, more than triple the size of the original OXE. We conduct a systematic study of how scaling robot augmentation impacts cross-embodiment learning. Results suggest that augmenting datasets with diverse arms and grippers improves policy performance not only on the augmented robots, but also on unseen robots and even the original robots under distribution shifts. In physical experiments, we demonstrate that state-of-the-art generalist policies such as OpenVLA and $π_0$ benefit from fine-tuning on OXE-AugE, improving success rates by 24-45% on previously unseen robot-gripper combinations across four real-world manipulation tasks. Project website: https://OXE-AugE.github.io/.

📄 PDF Abstract BibTeX arXiv:2512.13100

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AugESC: Dialogue Augmentation with Large Language Models for Emotional Support Conversation

2022-02-26 · Chujie Zheng, Sahand Sabour, Jiaxin Wen, Zheng Zhang 외

Crowdsourced dialogue corpora are usually limited in scale and topic coverage due to the expensive cost of data curation. This would hinder the generalization of downstream dialogue models to open-domain topics. In this …

Data AugmentationDialogue GenerationLanguage ModellingTopic coverage

Scale redundancy and soft gauge fixing in positively homogeneous neural networks

2026-02-16 · Rodrigo Carmo Terin arxiv

Neural networks with positively homogeneous activations exhibit an exact continuous reparametrization symmetry: neuron-wise rescalings generate parameter-space orbits along which the input--output function is invariant. …

Under pressure: learning-based analog gauge reading in the wild

2024-04-12 · Maurits Reitsma, Julian Keller, Kenneth Blomqvist, Roland Siegwart

We propose an interpretable framework for reading analog gauges that is deployable on real world robotic systems. Our framework splits the reading task into distinct steps, such that we can detect potential failures at e…

Generative Precipitation Downscaling using Score-based Diffusion with Wasserstein Regularization

2024-10-01 · Yuhao Liu, James Doss-Gollin, Guha Balakrishnan, Ashok Veeraraghavan

Understanding local risks from extreme rainfall, such as flooding, requires both long records (to sample rare events) and high-resolution products (to assess localized hazards). Unfortunately, there is a dearth of long-r…

Denoising

Machine learning for four-dimensional SU(3) lattice gauge theories

2026-04-14 · Urs Wenger arxiv

In this review I summarize how machine learning can be used in lattice gauge theory simulations and what ap\-proaches are currently available to improve the sampling of gauge field configurations, with a focus on applica…