Learning Condition Invariant Features for Retrieval-Based Localization from 1M Images
Image features for retrieval-based localization must be invariant to dynamic objects (e.g. cars) as well as seasonal and daytime changes. Such invariances are, up to some extent, learnable with existing methods using triplet-like losses, given a large number of diverse training images. However, due to the high algorithmic training complexity, there exists insufficient comparison between different loss functions on large datasets. In this paper, we train and evaluate several localization methods on three different benchmark datasets, including Oxford RobotCar with over one million images. This large scale evaluation yields valuable insights into the generalizability and performance of retrieval-based localization. Based on our findings, we develop a novel method for learning more accurate and better generalizing localization features. It consists of two main contributions: (i) a feature volume-based loss function, and (ii) hard positive and pairwise negative mining. On the challenging Oxford RobotCar night condition, our method outperforms the well-known triplet loss by 24.4% in localization accuracy within 5m.
Code (1)
Tasks
RetrievalTripletMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Retrieval-based Localization Based on Domain-invariant Feature Learning under Changing Environments
Visual localization is a crucial problem in mobile robotics and autonomous driving. One solution is to retrieve images with known pose from a database for the localization of query images. However, in environments with d…
Autonomous DrivingRetrievalTranslationVisual Localizationi3dLoc: Image-to-range Cross-domain Localization Robust to Inconsistent Environmental Conditions
We present a method for localizing a single camera with respect to a point cloud map in indoor and outdoor scenes. The problem is challenging because correspondences of local invariant features are inconsistent across th…
3D geometryGenerative Adversarial NetworkVisual LocalizationDomain-invariant Similarity Activation Map Contrastive Learning for Retrieval-based Long-term Visual Localization
Visual localization is a crucial component in the application of mobile robot and autonomous driving. Image retrieval is an efficient and effective technique in image-based localization methods. Due to the drastic variab…
Autonomous DrivingContrastive LearningImage-Based LocalizationImage Retrieval+4REST: Real-to-Synthetic Transform for Illumination Invariant Camera Localization
Accurate camera localization is an essential part of tracking systems. However, localization results are greatly affected by illumination. Including data collected under various lighting conditions can improve the robust…
Camera LocalizationGeometrically Mappable Image Features
Vision-based localization of an agent in a map is an important problem in robotics and computer vision. In that context, localization by learning matchable image features is gaining popularity due to recent advances in m…
Image RetrievalRetrieval