paper-with-me

홈 › Papers

FloorPlan-VLN: A New Paradigm for Floor Plan Guided Vision-Language Navigation

2026-03-18 · Kehan Chen, Yan Huang, Dong An, Jiawei He, Yifei Su, Jing Liu, Nianfeng Liu, Liang Wang arxiv

Existing Vision-Language Navigation (VLN) task requires agents to follow verbose instructions, ignoring some potentially useful global spatial priors, limiting their capability to reason about spatial structures. Although human-readable spatial schematics (e.g., floor plans) are ubiquitous in real-world buildings, current agents lack the cognitive ability to comprehend and utilize them. To bridge this gap, we introduce \textbf{FloorPlan-VLN}, a new paradigm that leverages structured semantic floor plans as global spatial priors to enable navigation with only concise instructions. We first construct the FloorPlan-VLN dataset, which comprises over 10k episodes across 72 scenes. It pairs more than 100 semantically annotated floor plans with Matterport3D-based navigation trajectories and concise instructions that omit step-by-step guidance. Then, we propose a simple yet effective method \textbf{FP-Nav} that uses a dual-view, spatio-temporally aligned video sequence, and auxiliary reasoning tasks to align observations, floor plans, and instructions. When evaluated under this new benchmark, our method significantly outperforms adapted state-of-the-art VLN baselines, achieving more than a 60\% relative improvement in navigation success rate. Furthermore, comprehensive noise modeling and real-world deployments demonstrate the feasibility and robustness of FP-Nav to actuation drift and floor plan distortions. These results validate the effectiveness of floor plan guided navigation and highlight FloorPlan-VLN as a promising step toward more spatially intelligent navigation.

📄 PDF Abstract BibTeX arXiv:2603.17437

Code (0)

등록된 구현이 없습니다.

Tasks

Vision-Language Navigation

Similar Papers 제목 키워드 기반

FloorplanVLM: A Vision-Language Model for Floorplan Vectorization

2026-02-06 · Yuanqing Liu, Ziming Yang, Yulong Li, Yue Yang arxiv

Converting raster floorplans into engineering-grade vector graphics is challenging due to complex topology and strict geometric constraints. To address this, we present FloorplanVLM, a unified framework that reformulates…

FloorplanMAE:A self-supervised framework for complete floorplan generation from partial inputs

2025-06-10 · Jun Yin, Jing Zhong, Pengyu Zeng, Peilin Li 외

In the architectural design process, floorplan design is often a dynamic and iterative process. Architects progressively draw various parts of the floorplan according to their ideas and requirements, continuously adjusti…

Self-Supervised Learning

Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation

2025-12-13 · Aydin Ayanzadeh, Tim Oates arxiv

Indoor navigation remains a critical challenge for people with visual impairments. The current solutions mainly rely on infrastructure-based systems, which limit their ability to navigate safely in dynamic environments. …

Zero-Shot LearningFew-Shot LearningKnowledge GraphsVisual Reasoning

End-to-end Graph-constrained Vectorized Floorplan Generation with Panoptic Refinement

2022-07-27 · Jiachen Liu, Yuan Xue, Jose Duarte, Krishnendra Shekhawat 외

The automatic generation of floorplans given user inputs has great potential in architectural design and has recently been explored in the computer vision community. However, the majority of existing methods synthesize f…

SLIBO-Net: Floorplan Reconstruction via Slicing Box Representation with Local Geometry Regularization

2023-09-21 · NeurIPS 2023 11

This paper focuses on improving the reconstruction of 2D floorplans from unstructured 3D point clouds. We identify opportunities for enhancement over the existing methods in three main areas: semantic quality, efficient …