paper-with-me

Papers

A Streamlined Attention-Based Network for Descriptor Extraction

2026-01-19 · Mattia D'Urso, Emanuele Santellani, Christian Sormann, Mattia Rossi, Andreas Kuhn, Friedrich Fraundorfer arxiv

We introduce SANDesc, a Streamlined Attention-Based Network for Descriptor extraction that aims to improve on existing architectures for keypoint description. Our descriptor network learns to compute descriptors that improve matching without modifying the underlying keypoint detector. We employ a revised U-Net-like architecture enhanced with Convolutional Block Attention Modules and residual paths, enabling effective local representation while maintaining computational efficiency. We refer to the building blocks of our model as Residual U-Net Blocks with Attention. The model is trained using a modified triplet loss in combination with a curriculum learning-inspired hard negative mining strategy, which improves training stability. Extensive experiments on HPatches, MegaDepth-1500, and the Image Matching Challenge 2021 show that training SANDesc on top of existing keypoint detectors leads to improved results on multiple matching tasks compared to the original keypoint descriptors. At the same time, SANDesc has a model complexity of just 2.4 million parameters. As a further contribution, we introduce a new urban dataset featuring 4K images and pre-calibrated intrinsics, designed to evaluate feature extractors. On this benchmark, SANDesc achieves substantial performance gains over the existing descriptors while operating with limited computational resources.

📄 PDF Abstract BibTeX arXiv:2601.13126

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyImage Matching

Similar Papers 제목 키워드 기반

3D Point Cloud Descriptors in Hand-crafted and Deep Learning Age: State-of-the-Art

2018-02-07 · Xian-Feng Han, Shi-Jie Sun, Xiang-Yu Song, Guo-Qiang Xiao

The introduction of inexpensive 3D data acquisition devices has promisingly facilitated the wide availability and popularity of 3D point cloud, which attracts more attention to the effective extraction of novel 3D point …

DRTAM: Dual Rank-1 Tensor Attention Module

2022-03-11 · Hanxing Chi, Baihong Lin, Jun Hu, Liang Wang

Recently, attention mechanisms have been extensively investigated in computer vision, but few of them show excellent performance on both large and mobile networks. This paper proposes Dual Rank-1 Tensor Attention Module …

IMFNet: Interpretable Multimodal Fusion for Point Cloud Registration

2021-11-18 · Xiaoshui Huang, Wentao Qu, Yifan Zuo, Yuming Fang 외

The existing state-of-the-art point descriptor relies on structure information only, which omit the texture information. However, texture information is crucial for our humans to distinguish a scene part. Moreover, the c…

Point Cloud Registration

The Influence of Streamlined Music on Cognition and Mood

2016-10-13

Recent advances in sound engineering have led to the development of so-called streamlined music designed to reduce exogenous attention and improve endogenous attention. Although anecdotal reports suggest that streamlined…

Form

A Streamlined Encoder/Decoder Architecture for Melody Extraction

2018-10-30 · Tsung-Han Hsieh, Li Su, Yi-Hsuan Yang

Melody extraction in polyphonic musical audio is important for music signal processing. In this paper, we propose a novel streamlined encoder/decoder network that is designed for the task. We make two technical contribut…

DecoderMelody Extraction