paper-with-me

홈 › Papers

Scene Retrieval for Contextual Visual Mapping

2021-02-25 · William H. B. Smith, Michael Milford, Klaus D. McDonald-Maier, Shoaib Ehsan

Visual navigation localizes a query place image against a reference database of place images, also known as a visual map'. Localization accuracy requirements for specific areas of the visual map, scene classes', vary according to the context of the environment and task. State-of-the-art visual mapping is unable to reflect these requirements by explicitly targetting scene classes for inclusion in the map. Four different scene classes, including pedestrian crossings and stations, are identified in each of the Nordland and St. Lucia datasets. Instead of re-training separate scene classifiers which struggle with these overlapping scene classes we make our first contribution: defining the problem of scene retrieval'. Scene retrieval extends image retrieval to classification of scenes defined at test time by associating a single query image to reference images of scene classes. Our second contribution is a triplet-trained convolutional neural network (CNN) to address this problem which increases scene classification accuracy by up to 7% against state-of-the-art networks pre-trained for scene recognition. The second contribution is an algorithm DMC' that combines our scene classification with distance and memorability for visual mapping. Our analysis shows that DMC includes 64% more images of our chosen scene classes in a visual map than just using distance interval mapping. State-of-the-art visual place descriptors AMOS-Net, Hybrid-Net and NetVLAD are finally used to show that DMC improves scene class localization accuracy by a mean of 3% and localization accuracy of the remaining map images by a mean of 10% across both datasets.

📄 PDF Abstract BibTeX arXiv:2102.12728

Code (0)

등록된 구현이 없습니다.

Tasks

General ClassificationImage RetrievalRetrievalScene ClassificationScene RecognitionTripletVisual Navigation

Similar Papers 제목 키워드 기반

Graph Similarities and Dual Approach for Sequential Text-to-Image Retrieval

2021-09-29 · Keonwoo Kim, Sihyeon Jo, Seong-Woo Kim

Sequential text-to-image retrieval, a.k.a. Story-to-images task, requires semantic alignment with a given story and maintaining global coherence in drawn image sequence simultaneously. Most of the previous works have onl…

Graph EmbeddingImage RetrievalRetrievalSentence+2

Beyond Visual Semantics: Exploring the Role of Scene Text in Image Understanding

2019-05-25 · Arka Ujjal Dey, Suman Kumar Ghosh, Ernest Valveny, Gaurav Harit

Images with visual and scene text content are ubiquitous in everyday life. However, current image interpretation systems are mostly limited to using only the visual features, neglecting to leverage the scene text content…

Retrieval

XeMap: Contextual Referring in Large-Scale Remote Sensing Environments

2025-04-30 · Yuxi Li, Lu Si, Yujie Hou, Chengaung Liu 외

Advancements in remote sensing (RS) imagery have provided high-resolution detail and vast coverage, yet existing methods, such as image-level captioning/retrieval and object-level detection/segmentation, often fail to ca…

What Images Cannot Say: Language-Guided Olfactory Representation Learning

2026-07-07 · Eleftherios Tsonis, Xi Wang, Vicky Kalogeiton arxiv

Images tell us what a scene looks like, but rarely what it would feel like to be there. While recent datasets pair visual scenes with electronic-nose measurements, aligning smell signals with images remains challenging b…

Representation LearningText Retrieval

Evaluation of Visual Place Recognition Methods for Image Pair Retrieval in 3D Vision and Robotics

2026-03-14 · Dennis Haitz, Athradi Shritish Shetty, Michael Weinmann, Markus Ulrich arxiv

Visual Place Recognition (VPR) is a core component in computer vision, typically formulated as an image retrieval task for localization, mapping, and navigation. In this work, we instead study VPR as an image pair retrie…

Visual Place RecognitionImage Retrieval