paper-with-me

Vision and Language Navigation

5개 벤치마크 · 논문 224편 · 이 태스크의 논문 보기 →

Benchmarks

VLN Challenge

결과 145개

Touchdown Dataset

결과 12개

RxR

결과 6개

map2seq

결과 5개

robo-vln

결과 1개

Most implemented

Papers

AeroDuo: Aerial Duo for UAV-based Vision and Language Navigation

2025-08-21 · Ruipu Wu, Yige Zhang, Jinyu Chen, Linjiang Huang 외 arxiv

Aerial Vision-and-Language Navigation (VLN) is an emerging task that enables Unmanned Aerial Vehicles (UAVs) to navigate outdoor environments using natural language instructions and visual cues. However, due to the exten…

Vision and Language Navigation

Rethinking the Embodied Gap in Vision-and-Language Navigation: A Holistic Study of Physical and Visual Disparities

2025-07-17 · Liuyi Wang, Xinyuan Xia, Hui Zhao, Hanqing Wang 외

Recent Vision-and-Language Navigation (VLN) advancements are promising, but their idealized assumptions about robot movement and control fail to reflect physically embodied deployment challenges. To bridge this gap, we i…

Large Language ModelVision and Language Navigation

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments

2025-06-30 · Xuan Yao, Junyu Gao, Changsheng Xu

Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to execute sequential navigation actions in complex environments guided by natural language instructions. Current approaches often strugg…

Decision MakingVision and Language Navigation

Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding

2025-06-12 · Yuhang Zhang, Haosheng Yu, Jiaping Xiao, Mir Feroskhan

Vision-and-language navigation (VLN) is a long-standing challenge in autonomous robotics, aiming to empower agents with the ability to follow human instructions while navigating complex environments. Two key bottlenecks …

Language ModelingLanguage ModellingLarge Language ModelVision and Language Navigation+1

A Navigation Framework Utilizing Vision-Language Models

2025-06-11 · Yicheng Duan, Kaiyu Tang

Vision-and-Language Navigation (VLN) presents a complex challenge in embodied AI, requiring agents to interpret natural language instructions and navigate through visually rich, unfamiliar environments. Recent advances i…

NavigatePrompt EngineeringVision and Language Navigation

Disrupting Vision-Language Model-Driven Navigation Services via Adversarial Object Fusion

2025-05-29 · Chunlong Xie, Jialing He, Shangwei Guo, Jiacheng Wang 외

We present Adversarial Object Fusion (AdvOF), a novel attack framework targeting vision-and-language navigation (VLN) agents in service-oriented environments by generating adversarial 3D objects. While foundational model…

Language ModelingLanguage ModellingObjectService Composition+1

전체 224편 보기 →