paper-with-me

홈 › Papers

AdaVLN: Towards Visual Language Navigation in Continuous Indoor Environments with Moving Humans

2024-11-27 · Dillon Loh, Tomasz Bednarz, Xinxing Xia, Frank Guan

Visual Language Navigation is a task that challenges robots to navigate in realistic environments based on natural language instructions. While previous research has largely focused on static settings, real-world navigation must often contend with dynamic human obstacles. Hence, we propose an extension to the task, termed Adaptive Visual Language Navigation (AdaVLN), which seeks to narrow this gap. AdaVLN requires robots to navigate complex 3D indoor environments populated with dynamically moving human obstacles, adding a layer of complexity to navigation tasks that mimic the real-world. To support exploration of this task, we also present AdaVLN simulator and AdaR2R datasets. The AdaVLN simulator enables easy inclusion of fully animated human models directly into common datasets like Matterport3D. We also introduce a "freeze-time" mechanism for both the navigation task and simulator, which pauses world state updates during agent inference, enabling fair comparisons and experimental reproducibility across different hardware. We evaluate several baseline models on this task, analyze the unique challenges introduced by AdaVLN, and demonstrate its potential to bridge the sim-to-real gap in VLN research.

📄 PDF Abstract BibTeX arXiv:2411.18539

Code (1)

dillonloh/adavln 공식 구현

Tasks

Navigate

Similar Papers 제목 키워드 기반

OptiSight: Bridging Semantic Reasoning and Geometric Control for Embodied Navigation

2026-08-24 · Alperen Avan, Jordi Sanchez-Riera arxiv

Autonomous indoor navigation requires both semantic understanding and precise geometric control. We propose OptiSight, a hybrid framework that combines Vision-Language Model reasoning with deterministic visual servoing t…

IndoorUAV: Benchmarking Vision-Language UAV Navigation in Continuous Indoor Environments

2025-12-22 · Xu Liu, Yu Liu, Hanshuo Qiu, Yang Qirong 외 arxiv

Vision-Language Navigation (VLN) enables agents to navigate in complex environments by following natural language instructions grounded in visual observations. Although most existing work has focused on ground-based robo…

Vision-Language NavigationMultimodal ReasoningData Augmentation

Towards blind user's indoor navigation: a comparative study of beacons and decawave for indoor accurate location

2019-12-02 · Prabin Sharma, Sambad Bidari, Kisan Thapa, Antonio Valente 외

There are many systems for indoor navigation specially built for visually impaired people but only some has good accuracy for navigation. While there are solutions like global navigation satellite systems for the localiz…

Indoor Localization

NavVerse: Benchmarking Indoor-to-Outdoor Embodied Navigation in Continuous Robot Simulation

2026-07-22 · Junzhe Wu, Yue Hu, Zeyu Han, Po-Hsun Chang 외 arxiv

Robots deployed in delivery, campus, and emergency-response settings often need to navigate from buildings to streets within a single continuous episode. Existing benchmarks usually evaluate indoor and outdoor navigation…

Fine-Tuning Vision-Language Models for Visual Navigation Assistance

2025-09-09 · Xiao Li, Bharat Gandhi, Ming Zhan, Mohit Nehra 외 arxiv

We address vision-language-driven indoor navigation to assist visually impaired individuals in reaching a target location using images and natural language guidance. Traditional navigation systems are ineffective indoors…

Visual Navigation