paper-with-me

Papers

Exploring Modality-shared Appearance Features and Modality-invariant Relation Features for Cross-modality Person Re-Identification

2021-04-23 · Nianchang Huang, Jianan Liu, Qiang Zhang, Jungong Han

Most existing cross-modality person re-identification works rely on discriminative modality-shared features for reducing cross-modality variations and intra-modality variations. Despite some initial success, such modality-shared appearance features cannot capture enough modality-invariant discriminative information due to a massive discrepancy between RGB and infrared images. To address this issue, on the top of appearance features, we further capture the modality-invariant relations among different person parts (referred to as modality-invariant relation features), which are the complement to those modality-shared appearance features and help to identify persons with similar appearances but different body shapes. To this end, a Multi-level Two-streamed Modality-shared Feature Extraction (MTMFE) sub-network is designed, where the modality-shared appearance features and modality-invariant relation features are first extracted in a shared 2D feature space and a shared 3D feature space, respectively. The two features are then fused into the final modality-shared features such that both cross-modality variations and intra-modality variations can be reduced. Besides, a novel cross-modality quadruplet loss is proposed to further reduce the cross-modality variations. Experimental results on several benchmark datasets demonstrate that our proposed method exceeds state-of-the-art algorithms by a noticeable margin.

📄 PDF Abstract BibTeX arXiv:2104.11539

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Modality Person Re-identificationPerson Re-Identification

Similar Papers 제목 키워드 기반

Challenge-Aware RGBT Tracking

2020-07-26 · ECCV 2020 8 · Chenglong Li, Lei Liu, Andong Lu, Qing Ji 외

RGB and thermal source data suffer from both shared and specific challenges, and how to explore and exploit them plays a critical role to represent the target appearance in RGBT tracking. In this paper, we propose a nove…

Rgb-T Tracking

RGBT Tracking via Multi-Adapter Network with Hierarchical Divergence Loss

2020-11-14 · Andong Lu, Chenglong Li, Yuqing Yan, Jin Tang 외

RGBT tracking has attracted increasing attention since RGB and thermal infrared data have strong complementary advantages, which could make trackers all-day and all-weather work. However, how to effectively represent RGB…

Representation LearningRgb-T TrackingVisual Tracking

Modality-Aware and Anatomical Vector-Quantized Autoencoding for Multimodal Brain MRI

2026-04-06 · Mingjie Li, Edward Kim, Yue Zhao, Ehsan Adeli 외 arxiv

Learning a robust Variational Autoencoder (VAE) is a fundamental step for many deep learning applications in medical image analysis, such as MRI synthesizes. Existing brain VAEs predominantly focus on single-modality dat…

Specificity-preserving RGB-D Saliency Detection

2021-08-18 · ICCV 2021 10 · Tao Zhou, Deng-Ping Fan, Geng Chen, Yi Zhou 외

Salient object detection (SOD) on RGB and depth images has attracted more and more research interests, due to its effectiveness and the fact that depth cues can now be conveniently captured. Existing RGB-D SOD models usu…

Decoderobject-detectionObject DetectionSaliency Detection+4

Cooperative Cross-Stream Network for Discriminative Action Representation

2019-08-27 · Jingran Zhang, Fumin Shen, Xing Xu, Heng Tao Shen

Spatial and temporal stream model has gained great success in video action recognition. Most existing works pay more attention to designing effective features fusion methods, which train the two-stream model in a separat…

Action RecognitionTemporal Action LocalizationTriplet