paper-with-me

Papers

Bidirectional Multi-Step Domain Generalization for Visible-Infrared Person Re-Identification

2024-03-16 · Mahdi Alehdaghi, Pourya Shamsolmoali, Rafael M. O. Cruz, Eric Granger

A key challenge in visible-infrared person re-identification (V-I ReID) is training a backbone model capable of effectively addressing the significant discrepancies across modalities. State-of-the-art methods that generate a single intermediate bridging domain are often less effective, as this generated domain may not adequately capture sufficient common discriminant information. This paper introduces the Bidirectional Multi-step Domain Generalization (BMDG), a novel approach for unifying feature representations across diverse modalities. BMDG creates multiple virtual intermediate domains by finding and aligning body part features extracted from both I and V modalities. Indeed, BMDG aims to reduce the modality gaps in two steps. First, it aligns modalities in feature space by learning shared and modality-invariant body part prototypes from V and I images. Then, it generalizes the feature representation by applying bidirectional multi-step learning, which progressively refines feature representations in each step and incorporates more prototypes from both modalities. In particular, our method minimizes the cross-modal gap by identifying and aligning shared prototypes that capture key discriminative features across modalities, then uses multiple bridging steps based on this information to enhance the feature representation. Experiments conducted on challenging V-I ReID datasets indicate that our BMDG approach outperforms state-of-the-art part-based models or methods that generate an intermediate domain from V-I person ReID.

📄 PDF Abstract BibTeX arXiv:2403.10782

Code (0)

등록된 구현이 없습니다.

Tasks

Domain GeneralizationPerson Re-Identification

Similar Papers 제목 키워드 기반

ML-BPM: Multi-teacher Learning with Bidirectional Photometric Mixing for Open Compound Domain Adaptation in Semantic Segmentation

2022-07-19 · Fei Pan, Sungsu Hur, Seokju Lee, Junsik Kim 외

Open compound domain adaptation (OCDA) considers the target domain as the compound of multiple unknown homogeneous subdomains. The goal of OCDA is to minimize the domain gap between the labeled source domain and the unla…

Domain AdaptationSemantic Segmentation

Generative Adversarial Network-based Synthesis of Visible Faces from Polarimetric Thermal Faces

2017-08-08 · He Zhang, Vishal M. Patel, Benjamin S. Riggan, Shuowen Hu

The large domain discrepancy between faces captured in polarimetric (or conventional) thermal and visible domain makes cross-domain face recognition quite a challenging problem for both human-examiners and computer visio…

Face GenerationFace RecognitionFace ReconstructionGenerative Adversarial Network+1

Induction in Both Directions: A Mechanistic Analysis of In-Context Learning in Masked Diffusion Language Models

2026-07-17 · Andy Catruna, Emilian Radoi arxiv

While the internal mechanisms of autoregressive (AR) transformers have been studied extensively, much less is known about diffusion language models (DLMs), an emerging alternative that generates text by iterative denoisi…

Template-based Multi-Domain Face Recognition

2024-09-15 · Anirudh Nanduri, Rama Chellappa

Despite the remarkable performance of deep neural networks for face detection and recognition tasks in the visible spectrum, their performance on more challenging non-visible domains is comparatively still lacking. While…

Domain AdaptationDomain GeneralizationFace DetectionFace Recognition

BiFM: Bidirectional Flow Matching for Few-Step Image Editing and Generation

2026-03-26 · Yasong Dai, Zeeshan Hayder, David Ahmedt-Aristizabal, Hongdong Li arxiv

Recent diffusion and flow matching models have demonstrated strong capabilities in image generation and editing by progressively removing noise through iterative sampling. While this enables flexible inversion for semant…

Image GenerationImage Editing