paper-with-me

홈 › Papers

ContextDesc: Local Descriptor Augmentation with Cross-Modality Context

2019-04-08 · CVPR 2019 6 · Zixin Luo, Tianwei Shen, Lei Zhou, Jiahui Zhang, Yao Yao, Shiwei Li, Tian Fang, Long Quan

Most existing studies on learning local features focus on the patch-based descriptions of individual keypoints, whereas neglecting the spatial relations established from their keypoint locations. In this paper, we go beyond the local detail representation by introducing context awareness to augment off-the-shelf local feature descriptors. Specifically, we propose a unified learning framework that leverages and aggregates the cross-modality contextual information, including (i) visual context from high-level image representation, and (ii) geometric context from 2D keypoint distribution. Moreover, we propose an effective N-pair loss that eschews the empirical hyper-parameter search and improves the convergence. The proposed augmentation scheme is lightweight compared with the raw local feature description, meanwhile improves remarkably on several large-scale benchmarks with diversified scenes, which demonstrates both strong practicality and generalization ability in geometric matching applications.

📄 PDF Abstract BibTeX arXiv:1904.04084

Code (1)

lzx551402/contextdesc 공식 구현 tf

Tasks

Geometric Matching

Similar Papers 제목 키워드 기반

Learning to Guide Local Feature Matches

2020-10-21 · François Darmon, Mathieu Aubry, Pascal Monasse

We tackle the problem of finding accurate and robust keypoint correspondences between images. We propose a learning-based approach to guide local feature matches via a learned approximate image matching. Our approach can…

3D Reconstruction

MatChA: Cross-Algorithm Matching with Feature Augmentation

2025-06-27 · Paula Carbó Cubero, Alberto Jaenal Gálvez, André Mateus, José Araújo 외

State-of-the-art methods fail to solve visual localization in scenarios where different devices use different sparse feature extraction algorithms to obtain keypoints and their corresponding descriptors. Translating feat…

Visual Localization

Descriptor-Injected Cross-Modal Learning: A Systematic Exploration of Audio-MIDI Alignment via Spectral and Melodic Features

2026-04-11 · Mariano Fernández Méndez arxiv

Cross-modal retrieval between audio recordings and symbolic music representations (MIDI) remains challenging because continuous waveforms and discrete event sequences encode different aspects of the same performance. We …

Cross-Modal Retrieval

LDCA: Local Descriptors with Contextual Augmentation for Few-Shot Learning

2024-01-24 · Maofa Wang, Bingchen Yan

Few-shot image classification has emerged as a key challenge in the field of computer vision, highlighting the capability to rapidly adapt to new tasks with minimal labeled data. Existing methods predominantly rely on im…

ClassificationFew-Shot Image ClassificationFew-Shot Learningimage-classification+1

RoRD: Rotation-Robust Descriptors and Orthographic Views for Local Feature Matching

2021-03-15 · Udit Singh Parihar, Aniket Gujarathi, Kinal Mehta, Satyajit Tourani 외

The use of local detectors and descriptors in typical computer vision pipelines work well until variations in viewpoint and appearance change become extreme. Past research in this area has typically focused on one of two…

Pose EstimationVisual Place Recognition