paper-with-me

Papers

GRM: Gradient Rectification Module for Visual Place Retrieval

2022-04-23 · Boshu Lei, Wenjie Ding, Limeng Qiao, Xi Qiu

Visual place retrieval aims to search images in the database that depict similar places as the query image. However, global descriptors encoded by the network usually fall into a low dimensional principal space, which is harmful to the retrieval performance. We first analyze the cause of this phenomenon, pointing out that it is due to degraded distribution of the gradients of descriptors. Then, we propose Gradient Rectification Module(GRM) to alleviate this issue. GRM is appended after the final pooling layer and can rectify gradients to the complementary space of the principal space. With GRM, the network is encouraged to generate descriptors more uniformly in the whole space. At last, we conduct experiments on multiple datasets and generalize our method to classification task under prototype learning framework.

📄 PDF Abstract BibTeX arXiv:2204.10972

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval

Similar Papers 제목 키워드 기반

Marior: Margin Removal and Iterative Content Rectification for Document Dewarping in the Wild

2022-07-23 · Jiaxin Zhang, Canjie Luo, Lianwen Jin, Fengjun Guo 외

Camera-captured document images usually suffer from perspective and geometric deformations. It is of great value to rectify them when considering poor visual aesthetics and the deteriorated performance of OCR systems. Re…

Optical Character Recognition (OCR)

DSR: Direct Self-rectification for Uncalibrated Dual-lens Cameras

2018-09-26 · Ruichao Xiao, Wenxiu Sun, Jiahao Pang, Qiong Yan 외

With the developments of dual-lens camera modules,depth information representing the third dimension of thecaptured scenes becomes available for smartphones. It isestimated by stereo matching algorithms, taking as input …

Stereo MatchingStereo Matching Hand

Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition

2026-06-30 · Zhiyao Shu, Jiacheng Yang, Yang Lu, Waishan Qiu 외 arxiv

Urban-scale Visual Place Recognition (VPR) aims to identify the geographic location of a query image by matching it against a geo-tagged database. While recent methods achieve impressive performance, they overlook a seri…

Visual Place Recognition

Identity-Decoupled Anonymization for Visual Evidence in Multi-modal Retrieval-Augmented Generation

2026-04-26 · Zehua Cheng, Wei Dai, Jiahao Sun arxiv

Multi-modal retrieval-augmented generation (MRAG) systems retrieve visual evidence from large image corpora to ground the responses of large multi-modal models, yet the retrieved images frequently contain human faces who…

Face Recognition

Reasoning Step-by-Step: Temporal Sentence Localization in Videos via Deep Rectification-Modulation Network

2020-12-01 · COLING 2020 8 · Daizong Liu, Xiaoye Qu, Jianfeng Dong, Pan Zhou

Temporal sentence localization in videos aims to ground the best matched segment in an untrimmed video according to a given sentence query. Previous works in this field mainly rely on attentional frameworks to align the …

Sentence