paper-with-me

홈 › Papers

NavOL: Navigation Policy with Online Imitation Learning

2026-05-12 · Xiaofei Wei, Chun Gu, Li Zhang arxiv

Learning robust navigation policies remains a core challenge in robotics. Offline imitation learning suffers from distribution shift and compounding errors at rollout, while reinforcement learning requires reward engineering and learns inefficiently. In this paper, we propose NavOL, an online imitation learning paradigm that interacts with a simulator and updates itself using expert demonstrations gathered online. Built upon a pretrained navigation diffusion policy that maps local observations to future waypoints, NavOL trains in a rollout update loop: during rollout, the policy acts in the simulator and queries a global planner which has privileged access to the global environment for the optimal path segment as ground truth trajectory labels; during update, the policy is trained on the online collected observation trajectory pairs. This online imitation loop removes the need for reward design, improves learning efficiency, and mitigates distribution shift by training on the policy own explored rollouts. Built on IsaacLab with fast, high-fidelity parallel rendering and domain randomization of camera pose and start-goal pairs, our system scales across 50 scenes on 8 RTX 4090 GPUs, collecting over 2,000 new trajectories per hour, each averaging more than 400 steps. We also introduce an indoor visual navigation benchmark with predefined start and goal positions for zero-shot generalization. Extensive evaluations on simulation benchmarks, including the NavDP benchmark and our proposed benchmark, as well as carefully designed real-world experiments, demonstrate the effectiveness of NavOL, showing consistent performance gains in online imitation learning.

📄 PDF Abstract BibTeX arXiv:2605.11762

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-shot GeneralizationReinforcement LearningVisual Navigation

Similar Papers 제목 키워드 기반

Integrating Offline Pre-Training with Online Fine-Tuning: A Reinforcement Learning Approach for Robot Social Navigation

2025-10-01 · Run Su, Hao Fu, Shuai Zhou, Yingao Fu arxiv

Offline reinforcement learning (RL) has emerged as a promising framework for addressing robot social navigation challenges. However, inherent uncertainties in pedestrian behavior and limited environmental interaction dur…

Reinforcement Learning

DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation

2025-08-13 · Haoxiang Shi, Xiang Deng, Zaijing Li, Gongwei Chen 외 arxiv

Vision-Language Navigation in Continuous Environments (VLN-CE) requires agents to follow natural language instructions through free-form 3D spaces. Existing VLN-CE approaches typically use a two-stage waypoint planning f…

Vision-Language NavigationSpatial Reasoning

DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric Voxelization

2023-04-30 · Yanpeng Zhao, Siyu Gao, Yunbo Wang, Xiaokang Yang

Unsupervised learning of object-centric representations in dynamic visual scenes is challenging. Unlike most previous approaches that learn to decompose 2D images, we present DynaVol, a 3D scene generative model that uni…

DecoderNeRFNeural RenderingNovel View Synthesis+3

Avoidance of Manual Labeling in Robotic Autonomous Navigation Through Multi-Sensory Semi-Supervised Learning

2017-09-22 · Junhong Xu, Shangyue Zhu, Hanqing Guo, Shaoen Wu

Imitation learning holds the promise to address challenging robotic tasks such as autonomous navigation. It however requires a human supervisor to oversee the training process and send correct control commands to robots …

Autonomous NavigationImitation LearningSensor Fusion

Optimal control of nonlinear systems with unsymmetrical input constraints and its applications to the UAV circumnavigation problem

2020-05-27 · Yangguang Yu, Xiangke Wang, Zhiyong Sun, Lincheng Shen

In this paper, a novel design scheme is introduced to solve the optimal control problem for nonlinear systems with unsymmetrical and state-dependent input constraints. By introducing an initial stabilizing control policy…