paper-with-me

홈 › Papers

Generalizable Domain Adaptation for Sim-and-Real Policy Co-Training

2025-09-23 · Shuo Cheng, Liqian Ma, Zhenyang Chen, Ajay Mandlekar, Caelan Garrett, Danfei Xu arxiv

Behavior cloning has shown promise for robot manipulation, but real-world demonstrations are costly to acquire at scale. While simulated data offers a scalable alternative, particularly with advances in automated demonstration generation, transferring policies to the real world is hampered by various simulation and real domain gaps. In this work, we propose a unified sim-and-real co-training framework for learning generalizable manipulation policies that primarily leverages simulation and only requires a few real-world demonstrations. Central to our approach is learning a domain-invariant, task-relevant feature space. Our key insight is that aligning the joint distributions of observations and their corresponding actions across domains provides a richer signal than aligning observations (marginals) alone. We achieve this by embedding an Optimal Transport (OT)-inspired loss within the co-training framework, and extend this to an Unbalanced OT framework to handle the imbalance between abundant simulation data and limited real-world examples. We validate our method on challenging manipulation tasks, showing it can leverage abundant simulation data to achieve up to a 30% improvement in the real-world success rate and even generalize to scenarios seen only in simulation. Project webpage: https://ot-sim2real.github.io/.

📄 PDF Abstract BibTeX arXiv:2509.18631

Code (0)

등록된 구현이 없습니다.

Tasks

Robot ManipulationDomain Adaptation

Similar Papers 제목 키워드 기반

Mind the Gap: Towards Generalizable Autonomous Penetration Testing via Domain Randomization and Meta-Reinforcement Learning

2024-12-05 · Shicheng Zhou, Jingju Liu, Yuliang Lu, Jiahai Yang 외

With increasing numbers of vulnerabilities exposed on the internet, autonomous penetration testing (pentesting) has emerged as a promising research area. Reinforcement learning (RL) is a natural fit for studying this top…

Large Language ModelMeta Reinforcement LearningReinforcement Learning (RL)

EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human Data

2025-09-23 · Ryan Punamiya, Dhruv Patel, Patcharapong Aphiwetsa, Pranav Kuppili 외 arxiv

Egocentric human experience data presents a vast resource for scaling up end-to-end imitation learning for robotic manipulation. However, significant domain gaps in visual appearance, sensor modalities, and kinematics be…

Domain Adaptation

LoopSR: Looping Sim-and-Real for Lifelong Policy Adaptation of Legged Robots

2024-09-26 · Peilin Wu, Weiji Xie, Jiahang Cao, Hang Lai 외

Reinforcement Learning (RL) has shown its remarkable and generalizable capability in legged locomotion through sim-to-real transfer. However, while adaptive methods like domain randomization are expected to make policy m…

Contrastive LearningDecoderReinforcement Learning (RL)

Causality Meets Locality: Provably Generalizable and Scalable Policy Learning for Networked Systems

2025-10-24 · Hao Liang, Shuqing Shi, Yudi Zhang, Biwei Huang 외 arxiv

Large-scale networked systems, such as traffic, power, and wireless grids, challenge reinforcement-learning agents with both scale and environment shifts. To address these challenges, we propose GSAC (Generalizable and S…

Representation LearningDomain Generalization

Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation

2024-05-02 · Homanga Bharadhwaj, Roozbeh Mottaghi, Abhinav Gupta, Shubham Tulsiani

We seek to learn a generalizable goal-conditioned policy that enables zero-shot robot manipulation: interacting with unseen objects in novel scenes without test-time adaptation. While typical approaches rely on a large a…

Robot ManipulationTest-time Adaptation