paper-with-me

홈 › Papers

Geometrically Mappable Image Features

2020-03-21 · Janine Thoma, Danda Pani Paudel, Ajad Chhatkuli, Luc van Gool

Vision-based localization of an agent in a map is an important problem in robotics and computer vision. In that context, localization by learning matchable image features is gaining popularity due to recent advances in machine learning. Features that uniquely describe the visual contents of images have a wide range of applications, including image retrieval and understanding. In this work, we propose a method that learns image features targeted for image-retrieval-based localization. Retrieval-based localization has several benefits, such as easy maintenance and quick computation. However, the state-of-the-art features only provide visual similarity scores which do not explicitly reveal the geometric distance between query and retrieved images. Knowing this distance is highly desirable for accurate localization, especially when the reference images are sparsely distributed in the scene. Therefore, we propose a novel loss function for learning image features which are both visually representative and geometrically relatable. This is achieved by guiding the learning process such that the feature and geometric distances between images are directly proportional. In our experiments we show that our features not only offer significantly better localization accuracy, but also allow to estimate the trajectory of a query sequence in absence of the reference images.

📄 PDF Abstract BibTeX arXiv:2003.09682

Code (1)

janinethoma/learning1M tf

Tasks

Image RetrievalRetrieval

Similar Papers 제목 키워드 기반

Approaching human 3D shape perception with neurally mappable models

2023-08-22 · Thomas P. O'Connell, Tyler Bonnen, Yoni Friedman, Ayush Tewari 외

Humans effortlessly infer the 3D shape of objects. What computations underlie this ability? Although various computational models have been proposed, none of them capture the human ability to match object shape across vi…

MULTI-VIEW LEARNING

Taxonomy grounded aggregation of classifiers with different label sets

2015-12-01 · Amrita Saha, Sathish Indurthi, Shantanu Godbole, Subendhu Rongali 외

We describe the problem of aggregating the label predictions of diverse classifiers using a class taxonomy. Such a taxonomy may not have been available or referenced when the individual classifiers were designed and trai…

General Classificationtext-classificationText Classification

Deformation-Invariant Neural Network and Its Applications in Distorted Image Restoration and Analysis

2023-10-04 · Han Zhang, Qiguang Chen, Lok Ming Lui

Images degraded by geometric distortions pose a significant challenge to imaging and computer vision tasks such as object recognition. Deep learning-based imaging models usually fail to give accurate performance for geom…

image-classificationImage ClassificationImage RestorationObject Recognition

Label-Attention Transformer with Geometrically Coherent Objects for Image Captioning

2021-09-16 · Shikha Dubey, Farrukh Olimov, Muhammad Aasim Rafique, Joonmo Kim 외

Automatic transcription of scene understanding in images and videos is a step towards artificial general intelligence. Image captioning is a nomenclature for describing meaningful information in an image using computer v…

DecoderImage CaptioningScene Understanding

LASER: LAtent SpacE Rendering for 2D Visual Localization

2022-04-01 · CVPR 2022 1 · Zhixiang Min, Naji Khosravan, Zachary Bessinger, Manjunath Narayana 외

We present LASER, an image-based Monte Carlo Localization (MCL) framework for 2D floor maps. LASER introduces the concept of latent space rendering, where 2D pose hypotheses on the floor map are directly rendered into a …

Indoor LocalizationMetric LearningVisual Localization