Localize Me Anywhere, Anytime: A Multi-Task Point-Retrieval Approach
Image-based localization is an essential complement to GPS localization. Current image-based localization methods are based on either 2D-to-3D or 3D-to-2D to find the correspondences, which ignore the real scene geometric attributes. The main contribution of our paper is that we use a 3D model reconstructed by a short video as the query to realize 3D-to-3D localization under a multi-task point retrieval framework. Firstly, the use of a 3D model as the query enables us to efficiently select location candidates. Furthermore, the reconstruction of 3D model exploits the correlations among different images, based on the fact that images captured from different views for SfM share information through matching features. By exploring shared information (matching features) across multiple related tasks (images of the same scene captured from different views), the visual feature's view-invariance property can be improved in order to get to a higher point retrieval accuracy. More specifically, we use multi-task point retrieval framework to explore the relationship between descriptors and the 3D points, which extracts the discriminant points for more accurate 3D-to-3D correspondences retrieval. We further apply multi-task learning (MTL) retrieval approach on thermal images to prove that our MTL retrieval framework also provides superior performance for the thermal domain. This application is exceptionally helpful to cope with the localization problem in an environment with limited light sources.
Code (0)
등록된 구현이 없습니다.
Tasks
Image-Based LocalizationMulti-Task LearningRetrievalSimilar Papers 제목 키워드 기반
A Korean Knowledge Extraction System for Enriching a KBox
The increased demand for structured knowledge has created considerable interest in knowledge extraction from natural language sentences. This study presents a new Korean knowledge extraction system and web interface for …
Entity LinkingRelation Extraction4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere
We present 4RC, a unified feed-forward framework for 4D reconstruction from monocular videos. Unlike existing approaches that typically decouple motion from geometry or produce limited 4D attributes such as sparse trajec…
Cognitive Systems Approach to Smart Cities
In our connected world, services are expected to be delivered at speed through multiple means with seamless communication. To put it in day to day conversational terms, 'there is an app for it' attitude prevails. Several…
SOON: Scenario Oriented Object Navigation with Graph-based Exploration
The ability to navigate like a human towards a language-guided target from anywhere in a 3D embodied environment is one of the 'holy grail' goals of intelligent robots. Most visual navigation benchmarks, however, focus o…
AttributeNavigateObjectVisual NavigationLearning-based Prediction and Uplink Retransmission for Wireless Virtual Reality (VR) Network
Wireless Virtual Reality (VR) users are able to enjoy immersive experience from anywhere at anytime. However, providing full spherical VR video with high quality under limited VR interaction latency is challenging. If th…