paper-with-me

홈 › Papers

X-Nav: Learning End-to-End Cross-Embodiment Navigation for Mobile Robots

2025-07-19 · Haitong Wang, Aaron Hao Tan, Angus Fung, Goldie Nejat arxiv

Existing navigation methods are primarily designed for specific robot embodiments, limiting their generalizability across diverse robot platforms. In this paper, we introduce X-Nav, a novel framework for end-to-end cross-embodiment navigation where a single unified policy can be deployed across various embodiments for both wheeled and quadrupedal robots. X-Nav consists of two learning stages: 1) multiple expert policies are trained using deep reinforcement learning with privileged observations on a wide range of randomly generated robot embodiments; and 2) a single general policy is distilled from the expert policies via navigation action chunking with transformer (Nav-ACT). The general policy directly maps visual and proprioceptive observations to low-level control commands, enabling generalization to novel robot embodiments. Simulated experiments demonstrated that X-Nav achieved zero-shot transfer to both unseen embodiments and photorealistic environments. A scalability study showed that the performance of X-Nav improves when trained with an increasing number of randomly generated embodiments. An ablation study confirmed the design choices of X-Nav. Furthermore, real-world experiments were conducted to validate the generalizability of X-Nav in real-world environments.

📄 PDF Abstract BibTeX arXiv:2507.14731

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

The One RING: a Robotic Indoor Navigation Generalist

2024-12-18 · Ainaz Eftekhar, Rose Hendrix, Luca Weihs, Jiafei Duan 외

Modern robots vary significantly in shape, size, and sensor configurations used to perceive and interact with their environments. However, most navigation policies are embodiment-specific--a policy trained on one robot t…

The Role of Embodiment in Intuitive Whole-Body Teleoperation for Mobile Manipulation

2025-09-03 · Sophia Bianchi Moyen, Rickmer Krohn, Sophie Lueth, Kay Pompetzki 외 arxiv

Intuitive Teleoperation interfaces are essential for mobile manipulation robots to ensure high quality data collection while reducing operator workload. A strong sense of embodiment combined with minimal physical and cog…

Mobile Robot Navigation Using Hand-Drawn Maps: A Vision Language Model Approach

2025-01-31 · Aaron Hao Tan, Angus Fung, Haitong Wang, Goldie Nejat

Hand-drawn maps can be used to convey navigation instructions between humans and robots in a natural and efficient manner. However, these maps can often contain inaccuracies such as scale distortions and missing landmark…

Language ModelingLanguage ModellingRobot Navigation

CeRLP: A Cross-embodiment Robot Local Planning Framework for Visual Navigation

2026-03-20 · Haoyu Xi, Mingao Tan, Xinming Zhang, Siwei Cheng 외 arxiv

Visual navigation for cross-embodiment robots is challenging due to variations in robot and camera configurations, which can lead to the failure of navigation tasks. Previous approaches typically rely on collecting massi…

Vision-Language NavigationMonocular Depth EstimationVisual Navigation

VAMOS: A Hierarchical Vision-Language-Action Model for Capability-Modulated and Steerable Navigation

2025-10-23 · Mateo Guaman Castro, Sidharth Rajagopal, Daniel Gorbatov, Matt Schmittle 외 arxiv

A fundamental challenge in robot navigation lies in learning policies that generalize across diverse environments while conforming to the unique physical constraints and capabilities of a specific embodiment (e.g., quadr…

Robot Navigation