paper-with-me

홈 › Papers

HA-VLN: A Benchmark for Human-Aware Navigation in Discrete-Continuous Environments with Dynamic Multi-Human Interactions, Real-World Validation, and an Open Leaderboard

2025-03-18 · Yifei Dong, Fengyi Wu, Qi He, Heng Li, Minghan Li, Zebang Cheng, Yuxuan Zhou, Jingdong Sun, Qi Dai, Zhi-Qi Cheng, Alexander G Hauptmann

Vision-and-Language Navigation (VLN) systems often focus on either discrete (panoramic) or continuous (free-motion) paradigms alone, overlooking the complexities of human-populated, dynamic environments. We introduce a unified Human-Aware VLN (HA-VLN) benchmark that merges these paradigms under explicit social-awareness constraints. Our contributions include: 1. A standardized task definition that balances discrete-continuous navigation with personal-space requirements; 2. An enhanced human motion dataset (HAPS 2.0) and upgraded simulators capturing realistic multi-human interactions, outdoor contexts, and refined motion-language alignment; 3. Extensive benchmarking on 16,844 human-centric instructions, revealing how multi-human dynamics and partial observability pose substantial challenges for leading VLN agents; 4. Real-world robot tests validating sim-to-real transfer in crowded indoor spaces; and 5. A public leaderboard supporting transparent comparisons across discrete and continuous tasks. Empirical results show improved navigation success and fewer collisions when social context is integrated, underscoring the need for human-centric design. By releasing all datasets, simulators, agent code, and evaluation tools, we aim to advance safer, more capable, and socially responsible VLN research.

📄 PDF Abstract BibTeX arXiv:2503.14229

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingHuman DynamicsVision and Language Navigation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

DynFly: Dynamic-Aware Continuous Trajectory Generation for UAV Vision-Language Navigation in Urban Environments

2026-06-30 · Wen Jiang, Hanfang Liang, Li Wang, Kangyao Huang 외 arxiv

Recent advances in multimodal large models have significantly improved UAV vision-language navigation (UAV-VLN) by enhancing high-level perception and reasoning. However, existing methods mainly focus on predicting discr…

Vision-Language Navigation

Bridging the Gap Between Learning in Discrete and Continuous Environments for Vision-and-Language Navigation

2022-03-05 · CVPR 2022 1 · Yicong Hong, Zun Wang, Qi Wu, Stephen Gould

Most existing works in vision-and-language navigation (VLN) focus on either discrete or continuous environments, training agents that cannot generalize across the two. The fundamental difference between the two setups is…

Imitation LearningVision and Language Navigation

Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments

2024-12-13 · Kehan Chen, Dong An, Yan Huang, Rongtao Xu 외

We address the task of Vision-Language Navigation in Continuous Environments (VLN-CE) under the zero-shot setting. Zero-shot VLN-CE is particularly challenging due to the absence of expert demonstrations for training and…

Vision-Language Navigation

Quantitative Metrics for Benchmarking Human-Aware Robot Navigation

2023-07-26 · IEEE Access 2023 7 · Jarosław Karwowski, Wojciech Szynkiewicz

Social robots have recently gained popularity, and many human-aware navigation approaches have emerged. This work presents a comprehensive benchmark for quantitatively assessing robot navigation methods. As an automated …

BenchmarkingRobot Navigation

Visual Representation Learning for Preference-Aware Path Planning

2021-09-18 · Kavan Singh Sikand, Sadegh Rabiee, Adam Uccello, Xuesu Xiao 외

Autonomous mobile robots deployed in outdoor environments must reason about different types of terrain for both safety (e.g., prefer dirt over mud) and deployer preferences (e.g., prefer dirt path over flower beds). Most…

Representation LearningSemantic Segmentation