paper-with-me

Papers

Frequency Domain Modality-invariant Feature Learning for Visible-infrared Person Re-Identification

2024-01-03 · Yulin Li, Tianzhu Zhang, Yongdong Zhang

Visible-infrared person re-identification (VI-ReID) is challenging due to the significant cross-modality discrepancies between visible and infrared images. While existing methods have focused on designing complex network architectures or using metric learning constraints to learn modality-invariant features, they often overlook which specific component of the image causes the modality discrepancy problem. In this paper, we first reveal that the difference in the amplitude component of visible and infrared images is the primary factor that causes the modality discrepancy and further propose a novel Frequency Domain modality-invariant feature learning framework (FDMNet) to reduce modality discrepancy from the frequency domain perspective. Our framework introduces two novel modules, namely the Instance-Adaptive Amplitude Filter (IAF) module and the Phrase-Preserving Normalization (PPNorm) module, to enhance the modality-invariant amplitude component and suppress the modality-specific component at both the image- and feature-levels. Extensive experimental results on two standard benchmarks, SYSU-MM01 and RegDB, demonstrate the superior performance of our FDMNet against state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2401.01839

Code (0)

등록된 구현이 없습니다.

Tasks

Metric LearningPerson Re-Identification

Similar Papers 제목 키워드 기반

WD-FQDet: Multispectral Detection Transformer via Wavelet Decomposition and Frequency-aware Query Learning

2026-05-13 · Chunjin Yang, Xiwei Zhang, Yiming Xiao, Fanman Meng arxiv

Infrared-visible object detection improves detection performance by combining complementary features from multispectral images. Existing backbone-specific and backbone-shared approaches still suffer from the problems of …

Object Detection

Frequency Domain Nuances Mining for Visible-Infrared Person Re-identification

2024-01-04 · Yukang Zhang, Yang Lu, Yan Yan, Hanzi Wang 외

The key of visible-infrared person re-identification (VIReID) lies in how to minimize the modality discrepancy between visible and infrared images. Existing methods mainly exploit the spatial information while ignoring t…

Face RecognitionPerson Re-Identification

ASFR-Net: Adversarial Alignment and Spatio-Frequency Refinement Network for Heterogeneous Remote Sensing Image Change Detection

2026-07-08 · Xin-Jie Wu, Zhi-Hui You, Si-Bao Chen, Qing-Ling Shu 외 arxiv

The core challenge of heterogeneous change detection in remote sensing imagery lies in effectively decoupling genuine land-cover changes from significant modal disparities caused by distinct imaging mechanisms. These int…

Change Detection

Learning Language-Driven Sequence-Level Modal-Invariant Representations for Video-Based Visible-Infrared Person Re-Identification

2026-01-17 · Xiaomei Yang, Antai Liu, Xizhan Gao, Fa Zhu 외 arxiv

The core of video-based visible-infrared person re-identification (VVI-ReID) lies in learning sequence-level modal-invariant representations across different modalities. Recent research tends to use modality-shared langu…

Person Re-IdentificationRepresentation Learning

Dual-Domain Perspective on Degradation-Aware Fusion: A VLM-Guided Robust Infrared and Visible Image Fusion Framework

2025-09-05 · Tianpei Zhang, Jufeng Zhao, Yiming Zhu, Guangmang Cui arxiv

Most existing infrared-visible image fusion (IVIF) methods assume high-quality inputs, and therefore struggle to handle dual-source degraded scenarios, typically requiring manual selection and sequential application of m…