Semantic Pose Verification for Outdoor Visual Localization with Self-supervised Contrastive Learning
Any city-scale visual localization system has to overcome long-term appearance changes, such as varying illumination conditions or seasonal changes between query and database images. Since semantic content is more robust to such changes, we exploit semantic information to improve visual localization. In our scenario, the database consists of gnomonic views generated from panoramic images (e.g. Google Street View) and query images are collected with a standard field-of-view camera at a different time. To improve localization, we check the semantic similarity between query and database images, which is not trivial since the position and viewpoint of the cameras do not exactly match. To learn similarity, we propose training a CNN in a self-supervised fashion with contrastive learning on a dataset of semantically segmented images. With experiments we showed that this semantic similarity estimation approach works better than measuring the similarity at pixel-level. Finally, we used the semantic similarity scores to verify the retrievals obtained by a state-of-the-art visual localization method and observed that contrastive learning-based pose verification increases top-1 recall value to 0.90 which corresponds to a 2% improvement.
Code (0)
등록된 구현이 없습니다.
Tasks
Contrastive LearningSemantic SimilaritySemantic Textual SimilarityVisual LocalizationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
MegLoc: A Robust and Accurate Visual Localization Pipeline
In this paper, we present a visual localization pipeline, namely MegLoc, for robust and accurate 6-DoF pose estimation under varying scenarios, including indoor and outdoor scenes, different time across a day, different …
Autonomous DrivingPose EstimationVisual LocalizationCrossLocate: Cross-modal Large-scale Visual Geo-Localization in Natural Environments using Rendered Modalities
We propose a novel approach to visual geo-localization in natural environments. This is a challenging problem due to vast localization areas, the variable appearance of outdoor environments and the scarcity of available …
Camera LocalizationCamera Pose Estimationgeo-localizationImage-Based Localization+4Retrieval and Localization with Observation Constraints
Accurate visual re-localization is very critical to many artificial intelligence applications, such as augmented reality, virtual reality, robotics and autonomous driving. To accomplish this task, we propose an integrate…
Autonomous DrivingImage RetrievalRetrievalSemantic SegmentationVisual-Inertial SLAM for Unstructured Outdoor Environments: Benchmarking the Benefits and Computational Costs of Loop Closing
Simultaneous Localization and Mapping (SLAM) is essential for mobile robotics, enabling autonomous navigation in dynamic, unstructured outdoor environments without relying on external positioning systems. These environme…
Autonomous NavigationBenchmarkingSimultaneous Localization and MappingIs This The Right Place? Geometric-Semantic Pose Verification for Indoor Visual Localization
Visual localization in large and complex indoor scenes, dominated by weakly textured rooms and repeating geometric patterns, is a challenging problem with high practical relevance for applications such as Augmented Reali…
Visual Localization