paper-with-me

홈 › Papers

A2GC: Asymmetric Aggregation with Geometric Constraints for Locally Aggregated Descriptors

2025-11-18 · Zhenyu Li, Tianyi Shang arxiv

Visual Place Recognition (VPR) aims to match query images against a database using visual cues. State-of-the-art methods aggregate features from deep backbones to form global descriptors. Optimal transport-based aggregation methods reformulate feature-to-cluster assignment as a transport problem, but the standard Sinkhorn algorithm symmetrically treats source and target marginals, limiting effectiveness when image features and cluster centers exhibit substantially different distributions. We propose an asymmetric aggregation VPR method with geometric constraints for locally aggregated descriptors, called $A^2$GC-VPR. Our method employs row-column normalization averaging with separate marginal calibration, enabling asymmetric matching that adapts to distributional discrepancies in visual place recognition. Geometric constraints are incorporated through learnable coordinate embeddings, computing compatibility scores fused with feature similarities, thereby promoting spatially proximal features to the same cluster and enhancing spatial awareness. Experimental results on MSLS, NordLand, and Pittsburgh datasets demonstrate superior performance, validating the effectiveness of our approach in improving matching accuracy and robustness.

📄 PDF Abstract BibTeX arXiv:2511.14109

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Place Recognition

Similar Papers 제목 키워드 기반

Tree-based Node Aggregation in Sparse Graphical Models

2021-01-29 · Ines Wilms, Jacob Bien

High-dimensional graphical models are often estimated using regularization that is aimed at reducing the number of edges in a network. In this work, we show how even simpler networks can be produced by aggregating the no…

TAG

Vectors of Locally Aggregated Centers for Compact Video Representation

2015-09-13 · Alhabib Abbas, Nikos Deligiannis, Yiannis Andreopoulos

We propose a novel vector aggregation technique for compact video representation, with application in accurate similarity detection within large video datasets. The current state-of-the-art in visual search is formed by …

ClusteringVideo Description

GoMVS: Geometrically Consistent Cost Aggregation for Multi-View Stereo

2024-04-11 · CVPR 2024 1 · Jiang Wu, Rui Li, Haofei Xu, Wenxun Zhao 외

Matching cost aggregation plays a fundamental role in learning-based multi-view stereo networks. However, directly aggregating adjacent costs can lead to suboptimal results due to local geometric inconsistency. Related m…

3D Reconstruction

Learning Decentralized Wireless Resource Allocations with Graph Neural Networks

2021-07-03 · Zhiyang Wang, Mark Eisen, Alejandro Ribeiro

We consider the broad class of decentralized optimal resource allocation problems in wireless networks, which can be formulated as a constrained statistical learning problems with a localized information structure. We de…

MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition

2026-05-10 · Zhengyi Xu, Yuhang Ming, Zhihao Zhan, Hanyu Zhu 외 arxiv

Multi-modal cross-view place recognition remains a fundamental challenge in computer vision and robotics due to the severe viewpoint, modality, and spatial-structure discrepancies between ground observations and aerial r…