paper-with-me

Papers

Learning a Unified Latent Space for Cross-Embodiment Robot Control

2026-01-21 · Yashuai Yan, Dongheui Lee arxiv

We present a scalable framework for cross-embodiment humanoid robot control by learning a shared latent representation that unifies motion across humans and diverse humanoid platforms, including single-arm, dual-arm, and legged humanoid robots. Our method proceeds in two stages: first, we construct a decoupled latent space that captures localized motion patterns across different body parts using contrastive learning, enabling accurate and flexible motion retargeting even across robots with diverse morphologies. To enhance alignment between embodiments, we introduce tailored similarity metrics that combine joint rotation and end-effector positioning for critical segments, such as arms. Then, we train a goal-conditioned control policy directly within this latent space using only human data. Leveraging a conditional variational autoencoder, our policy learns to predict latent space displacements guided by intended goal directions. We show that the trained policy can be directly deployed on multiple robots without any adaptation. Furthermore, our method supports the efficient addition of new robots to the latent space by learning only a lightweight, robot-specific embedding layer. The learned latent policies can also be directly applied to the new robots. Experimental results demonstrate that our approach enables robust, scalable, and embodiment-agnostic robot control across a wide range of humanoid platforms.

📄 PDF Abstract BibTeX arXiv:2601.15419

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Similar Papers 제목 키워드 기반

Cross-Hand Latent Representation for Vision-Language-Action Models

2026-03-10 · Guangqi Jiang, Yutong Liang, Jianglong Ye, Jia-Yang Huang 외 arxiv

Dexterous manipulation is essential for real-world robot autonomy, mirroring the central role of human hand coordination in daily activity. Humans rely on rich multimodal perception--vision, sound, and language-guided in…

PHASOR: Phase-Anchored Universal Action Representations for Humanoid Embodiments

2026-06-01 · Kihyun Kim, Chaeyun Kim, Jongho Shin, Taeyoun Kwon 외 arxiv

Learning a good action embedding space is fundamental to scalable robot policy learning, yet existing methods treat action latents as task-specific intermediates rather than first-class representations. The resulting lat…

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space

2026-07-03 · Luis Felipe Casas, Robert Teal, Keval Shah, Abhijit Tadepalli 외 arxiv

Robot manipulation policies are typically tied to specific robotic hand embodiments, limiting the transfer of learned behaviors across platforms with different kinematic structures. In this work, we propose the Unified H…

Reinforcement LearningRobot Manipulation

One-Policy-Fits-All: Geometry-Aware Action Latents for Cross-Embodiment Manipulation

2026-03-15 · Juncheng Mu, Sizhe Yang, Hojin Bae, Feiyu Jia 외 arxiv

Cross-embodiment manipulation is crucial for enhancing the scalability of robot manipulation and reducing the high cost of data collection. However, the significant differences between embodiments, such as variations in …

Robot Manipulation

SCAR: Self-Supervised Continuous Action Representation Learning

2026-05-13 · Hongjia Liu, Fan Feng, Minghao Fu, Xinyue Wang 외 arxiv

Despite the central role of action in embodied intelligence, learning transferable action representations from visual transitions remains a fundamental challenge, particularly when world models must generalize across emb…

Representation Learning