paper-with-me

Papers

Embedding Morphology into Transformers for Cross-Robot Policy Learning

2026-02-26 · Kei Suzuki, Jing Liu, Ye Wang, Chiori Hori, Matthew Brand, Diego Romeres, Toshiaki Koike-Akino arxiv

Cross-robot policy learning -- training a single policy to perform well across multiple embodiments -- remains a central challenge in robot learning. Transformer-based policies, such as vision-language-action (VLA) models, are typically embodiment-agnostic and must infer kinematic structure purely from observations, which can reduce robustness across embodiments and even limit performance within a single embodiment. We propose an embodiment-aware transformer policy that injects morphology via three mechanisms: (1) kinematic tokens that factorize actions across joints and compress time through per-joint temporal chunking; (2) a topology-aware attention bias that encodes kinematic topology as an inductive bias in self-attention, encouraging message passing along kinematic edges; and (3) joint-attribute conditioning that augments topology with per-joint descriptors to capture semantics beyond connectivity. Across a range of embodiments, this structured integration consistently improves performance over a vanilla pi0.5 VLA baseline, indicating improved robustness both within an embodiment and across embodiments.

📄 PDF Abstract BibTeX arXiv:2603.00182

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Universal Morphology Control via Contextual Modulation

2023-02-22 · Zheng Xiong, Jacob Beck, Shimon Whiteson

Learning a universal policy across different robot morphologies can significantly improve learning efficiency and generalization in continuous control. However, it poses a challenging multi-task reinforcement learning pr…

continuous-controlContinuous Control

Morpho-Aware Global Attention for Image Matting

2024-11-15 · Jingru Yang, Chengzhi Cao, Chentianye Xu, Zhongwei Xie 외

Vision Transformers (ViTs) and Convolutional Neural Networks (CNNs) face inherent challenges in image matting, particularly in preserving fine structural details. ViTs, with their global receptive field enabled by the se…

Image Matting

Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control

2024-02-09 · Zheng Xiong, Risto Vuorio, Jacob Beck, Matthieu Zimmer 외

Learning a universal policy across different robot morphologies can significantly improve learning efficiency and enable zero-shot generalization to unseen morphologies. However, learning a highly performant universal po…

Zero-shot Generalization

MetaMorph: Learning Universal Controllers with Transformers

2022-03-22 · ICLR 2022 4 · Agrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-Fei

Multiple domains like vision, natural language, and audio are witnessing tremendous progress by leveraging Transformers for large scale pre-training followed by task specific fine tuning. In contrast, in robotics we prim…

Zero-shot Generalization

Morphology-Conditioned World Model for Cross-Embodiment Quadrupedal Locomotion

2026-04-09 · Mohamad H. Danesh, Chenhao Li, Amin Abyaneh, Anas Houssaini 외 arxiv

World models promise a paradigm shift in robotics, where an agent learns the physics of its environment once and then acquires behaviors efficiently. Yet the learned dynamics models at their core are typically morphology…

Zero-shot Generalization