paper-with-me

홈 › Papers

Less Is More: Scalable Visual Navigation from Limited Data

2026-01-25 · Yves Inglin, Jonas Frey, Changan Chen, Marco Hutter arxiv

Imitation learning provides a powerful framework for goal-conditioned visual navigation in mobile robots, enabling obstacle avoidance while respecting human preferences and social norms. However, its effectiveness depends critically on the quality and diversity of training data. In this work, we show how classical geometric planners can be leveraged to generate synthetic trajectories that complement costly human demonstrations. We train Less is More (LiMo), a transformer-based visual navigation policy that predicts goal-conditioned SE(2) trajectories from a single RGB observation, and find that augmenting limited expert demonstrations with planner-generated supervision yields substantial performance gains. Through ablations and complementary qualitative and quantitative analyses, we characterize how dataset scale and diversity affect planning performance. We demonstrate real-robot deployment and argue that robust visual navigation is enabled not by simply collecting more demonstrations, but by strategically curating diverse, high-quality datasets. Our results suggest that scalable, embodiment-specific geometric supervision is a practical path toward data-efficient visual navigation.

📄 PDF Abstract BibTeX arXiv:2601.17815

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Navigation

Similar Papers 제목 키워드 기반

CREStE: Scalable Mapless Navigation with Internet Scale Priors and Counterfactual Guidance

2025-03-05 · Arthur Zhang, Harshit Sikchi, Amy Zhang, Joydeep Biswas

We introduce CREStE, a scalable learning-based mapless navigation framework to address the open-world generalization and robustness challenges of outdoor urban navigation. Key to achieving this is learning perceptual rep…

Active LearningcounterfactualNavigate

ImagiNav: Scalable Embodied Navigation via Generative Visual Prediction and Inverse Dynamics

2026-03-14 · Jie Chen, Yuxin Cai, Yizhuo Wang, Ruofei Bai 외 arxiv

Enabling robots to navigate open-world environments via natural language is critical for general-purpose autonomy. Yet, Vision-Language Navigation has relied on end-to-end policies trained on expensive, embodiment-specif…

Vision-Language NavigationRobot Navigation

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

2026-05-07 · Oğuzhan Fatih Kar, Roman Bachmann, Yuanzheng Gong, Anders Boesen Lindbo Larsen 외 arxiv

The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection attempts remain limited to offline trajectories for supervised fine-tu…

VLD: Visual Language Goal Distance for Reinforcement Learning Navigation

2025-12-08 · Lazar Milikic, Manthan Patel, Jonas Frey arxiv

Training end-to-end policies from image data to directly predict navigation actions for robotic systems has proven inherently difficult. Existing approaches often suffer from either the sim-to-real gap during policy tran…

Reinforcement Learning

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

2026-07-07 · Zihan Wang, Seungjun Lee, Yinghao Xu, Gim Hee Lee arxiv

Embodied navigation aims to build agents that interpret multimodal goals, reason in 3D space, and reach target destinations reliably in the real world. However, progress remains constrained by the lack of scalable, high-…