From Coarse to Fine: Robust Hierarchical Localization at Large Scale
Robust and accurate visual localization is a fundamental capability for numerous applications, such as autonomous driving, mobile robotics, or augmented reality. It remains, however, a challenging task, particularly for large-scale environments and in presence of significant appearance changes. State-of-the-art methods not only struggle with such scenarios, but are often too resource intensive for certain real-time applications. In this paper we propose HF-Net, a hierarchical localization approach based on a monolithic CNN that simultaneously predicts local features and global descriptors for accurate 6-DoF localization. We exploit the coarse-to-fine localization paradigm: we first perform a global retrieval to obtain location hypotheses and only later match local features within those candidate places. This hierarchical approach incurs significant runtime savings and makes our system suitable for real-time operation. By leveraging learned descriptors, our method achieves remarkable localization robustness across large variations of appearance and sets a new state-of-the-art on two challenging benchmarks for large-scale localization.
Code (3)
Tasks
Autonomous DrivingRetrievalVisual LocalizationVisual Place RecognitionSimilar Papers 제목 키워드 기반
Efficient Large-Scale Localization by Global Instance Recognition
Hierarchical frameworks consisting of both coarse and fine localization are often used as the standard pipeline for large-scale visual localization. Despite their promising performance in simple environments, they st…
Visual LocalizationProbabilistic Visual Place Recognition for Hierarchical Localization
Visual localization techniques often comprise a hierarchical localization pipeline, with a visual place recognition module used as a coarse localizer to initialize a pose refinement stage. While improving the pose refine…
Image RetrievalRetrievalState EstimationVisual Localization+1HSCNet++: Hierarchical Scene Coordinate Classification and Regression for Visual Localization with Transformer
Visual localization is critical to many applications in computer vision and robotics. To address single-image RGB localization, state-of-the-art feature-based methods match local descriptors between a query image and a p…
regressionVisual LocalizationLiteVLoc: Map-Lite Visual Localization for Image Goal Navigation
This paper presents LiteVLoc, a hierarchical visual localization framework that uses a lightweight topo-metric map to represent the environment. The method consists of three sequential modules that estimate camera poses …
Pose EstimationVisual LocalizationiSimLoc: Visual Global Localization for Previously Unseen Environments with Simulated Images
The visual camera is an attractive device in beyond visual line of sight (B-VLOS) drone operation, since they are low in size, weight, power, and cost, and can provide redundant modality to GPS failures. However, state-o…
RetrievalVisual Localization