paper-with-me

홈 › Papers

Fidelity-Aware Data Composition for Robust Robot Generalization

2025-09-29 · Zizhao Tong, Di Chen, Sicheng Hu, Hongwei Fan, Liliang Chen, Guanghui Ren, Hao Tang, Hao Dong, Ling Shao arxiv

Generalist robot policies trained on large-scale, visually homogeneous datasets can be susceptible to shortcut learning, which impairs their out-of-distribution (OOD) generalization. While generative data augmentation is a common approach to introduce diversity, it presents a subtle challenge: data composition. Naively mixing real and synthetic data can corrupt the learning signal, as this process often prioritizes visual diversity at the expense of information fidelity. This paper suggests that robust generalization depends on principled, fidelity-aware data composition. We introduce Coherent Information Fidelity Tuning (CIFT), a framework that treats data composition as an optimization problem. CIFT uses a practical proxy for Information Fidelity based on the feature-space geometry of a dataset. This enables the identification of a phase transition, termed the Decoherence Point, where training stability degrades. The framework includes a generative engine, Multi-View Video Augmentation (MVAug), to synthesize a causally disentangled data spectrum for this tuning process. Applying CIFT to policy architectures such as $π_0$ and Diffusion Policy improves OOD success rates by over 54\%. These results indicate that fidelity-aware composition, beyond data synthesis alone, is an important component for developing robust, general-purpose robots.

📄 PDF Abstract BibTeX arXiv:2509.24797

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling

2026-05-03 · Peiyan Tu, Hanxin Zhu, Jingwen Sun, Shaojie Ren 외 arxiv

Embodied agents require robust and comprehensive 3D spatiotemporal representations to support spatial reasoning, manipulation understanding, and downstream decision making. However, existing robot data are typically capt…

Spatial ReasoningDecision Making

Skill-Aware Diffusion for Generalizable Robotic Manipulation

2026-01-16 · Aoshen Huang, Jiaming Chen, Jiyu Cheng, Ran Song 외 arxiv

Robust generalization in robotic manipulation is crucial for robots to adapt flexibly to diverse environments. Existing methods usually improve generalization by scaling data and networks, but model tasks independently a…

Towards Generalizable Robotic Data Flywheel: High-Dimensional Factorization and Composition

2026-03-26 · Yuyang Xiao, Yifei Zhou, Haoran Wang, Wenxuan Ou 외 arxiv

The lack of sufficiently diverse data, coupled with limited data efficiency, remains a major bottleneck for generalist robotic models, yet systematic strategies for collecting and curating such data are not fully explore…

Scale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic Manipulation

2026-07-23 · Yu Qi, Zhang Ye, Xinyi Xu, Yuxuan Lu 외 arxiv

Compositional generalization is essential for robot to follow diverse instructions. However, pretrained policies are known to take shortcuts, deferring to salient cues rather than grounding language. We introduce a diagn…

A Multifidelity Sim-to-Real Pipeline for Verifiable and Compositional Reinforcement Learning

2023-12-02 · Cyrus Neary, Christian Ellis, Aryaman Singh Samyal, Craig Lennon 외

We propose and demonstrate a compositional framework for training and verifying reinforcement learning (RL) systems within a multifidelity sim-to-real pipeline, in order to deploy reliable and adaptable RL policies on ph…

reinforcement-learningReinforcement Learning (RL)