paper-with-me

홈 › Papers

MSCMNet: Multi-scale Semantic Correlation Mining for Visible-Infrared Person Re-Identification

2023-11-24 · Xuecheng Hua, Ke Cheng, Hu Lu, Juanjuan Tu, Yuanquan Wang, Shitong Wang

The main challenge in the Visible-Infrared Person Re-Identification (VI-ReID) task lies in how to extract discriminative features from different modalities for matching purposes. While the existing well works primarily focus on minimizing the modal discrepancies, the modality information can not thoroughly be leveraged. To solve this problem, a Multi-scale Semantic Correlation Mining network (MSCMNet) is proposed to comprehensively exploit semantic features at multiple scales and simultaneously reduce modality information loss as small as possible in feature extraction. The proposed network contains three novel components. Firstly, after taking into account the effective utilization of modality information, the Multi-scale Information Correlation Mining Block (MIMB) is designed to explore semantic correlations across multiple scales. Secondly, in order to enrich the semantic information that MIMB can utilize, a quadruple-stream feature extractor (QFE) with non-shared parameters is specifically designed to extract information from different dimensions of the dataset. Finally, the Quadruple Center Triplet Loss (QCT) is further proposed to address the information discrepancy in the comprehensive features. Extensive experiments on the SYSU-MM01, RegDB, and LLCM datasets demonstrate that the proposed MSCMNet achieves the greatest accuracy.

📄 PDF Abstract BibTeX arXiv:2311.14395

Code (1)

Hua-XC/MSCMNet 공식 구현 pytorch

Tasks

Person Re-IdentificationTriplet

Methods 이 논문이 사용한 방법론

Focus 설명 없음
Triplet Loss The goal of Triplet loss, in the context of Siamese Networks, is to maximize the joint probability among all score-pairs i.e. the product of all probabilities. By using its…

Similar Papers 제목 키워드 기반

Efficient Semantic Matching with Hypercolumn Correlation

2023-11-07 · SeungWook Kim, Juhong Min, Minsu Cho

Recent studies show that leveraging the match-wise relationships within the 4D correlation map yields significant improvements in establishing semantic correspondences - but at the cost of increased computation and laten…

Mining Relations among Cross-Frame Affinities for Video Semantic Segmentation

2022-07-21 · Guolei Sun, Yun Liu, Hao Tang, Ajad Chhatkuli 외

The essence of video semantic segmentation (VSS) is how to leverage temporal information for prediction. Previous efforts are mainly devoted to developing new techniques to calculate the cross-frame affinities such as op…

Optical Flow EstimationSemantic SegmentationVideo Semantic Segmentation

Adaptive Structural Similarity Preserving for Unsupervised Cross Modal Hashing

2022-07-09 · Liang Li, Baihua Zheng, Weiwei Sun

Cross-modal hashing is an important approach for multimodal data management and application. Existing unsupervised cross-modal hashing algorithms mainly rely on data features in pre-trained models to mine their similarit…

ManagementRepresentation Learning

Knowledge-aware Alert Aggregation in Large-scale Cloud Systems: a Hybrid Approach

2024-03-11 · Jinxi Kuang, Jinyang Liu, JunJie Huang, Renyi Zhong 외

Due to the scale and complexity of cloud systems, a system failure would trigger an "alert storm", i.e., massive correlated alerts. Although these alerts can be traced back to a few root causes, the overwhelming number m…

CoLALanguage ModellingLarge Language ModelSemantic Similarity+1

TransFusion: Multi-view Divergent Fusion for Medical Image Segmentation with Transformers

2022-03-21 · Di Liu, Yunhe Gao, Qilong Zhangli, Ligong Han 외

Combining information from multi-view images is crucial to improve the performance and robustness of automated methods for disease diagnosis. However, due to the non-alignment characteristics of multi-view images, buildi…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation