AmsterTime: A Visual Place Recognition Benchmark Dataset for Severe Domain Shift
We introduce AmsterTime: a challenging dataset to benchmark visual place recognition (VPR) in presence of a severe domain shift. AmsterTime offers a collection of 2,500 well-curated images matching the same scene from a street view matched to historical archival image data from Amsterdam city. The image pairs capture the same place with different cameras, viewpoints, and appearances. Unlike existing benchmark datasets, AmsterTime is directly crowdsourced in a GIS navigation platform (Mapillary). We evaluate various baselines, including non-learning, supervised and self-supervised methods, pre-trained on different relevant datasets, for both verification and retrieval tasks. Our result credits the best accuracy to the ResNet-101 model pre-trained on the Landmarks dataset for both verification and retrieval tasks by 84% and 24%, respectively. Additionally, a subset of Amsterdam landmarks is collected for feature evaluation in a classification task. Classification labels are further used to extract the visual explanations using Grad-CAM for inspection of the learned similar visuals in a deep metric learning models.
Code (1)
Tasks
Image ClassificationImage RetrievalMetric LearningRetrievalVisual Place RecognitionSimilar Papers 제목 키워드 기반
DC-VLAQ: Query-Residual Aggregation for Robust Visual Place Recognition
One of the central challenges in visual place recognition (VPR) is learning a robust global representation that remains discriminative under large viewpoint changes, illumination variations, and severe domain shifts. Whi…
Visual Place RecognitionA Fast and Robust Place Recognition Approach for Stereo Visual Odometry Using LiDAR Descriptors
Place recognition is a core component of Simultaneous Localization and Mapping (SLAM) algorithms. Particularly in visual SLAM systems, previously-visited places are recognized by measuring the appearance similarity betwe…
Computational EfficiencySimultaneous Localization and MappingVisual OdometrySpatio-Semantic ConvNet-Based Visual Place Recognition
We present a Visual Place Recognition system that follows the two-stage format common to image retrieval pipelines. The system encodes images of places by employing the activations of different layers of a pre-trained, o…
Image RetrievalRetrievalVisual Place RecognitionDeep Learning Features at Scale for Visual Place Recognition
The success of deep learning techniques in the computer vision domain has triggered a range of initial investigations into their utility for visual place recognition, all using generic features from networks that were tr…
Deep LearningVisual Place RecognitionPlaceFormer: Transformer-based Visual Place Recognition using Multi-Scale Patch Selection and Fusion
Visual place recognition is a challenging task in the field of computer vision, and autonomous robotics and vehicles, which aims to identify a location or a place from visual inputs. Contemporary methods in visual place …
Computational EfficiencyImage RetrievalVisual Place Recognition