paper-with-me

Papers

Asymmetric Modality Translation For Face Presentation Attack Detection

2021-10-18 · Zhi Li, Haoliang Li, Xin Luo, Yongjian Hu, Kwok-Yan Lam, Alex C. Kot

Face presentation attack detection (PAD) is an essential measure to protect face recognition systems from being spoofed by malicious users and has attracted great attention from both academia and industry. Although most of the existing methods can achieve desired performance to some extent, the generalization issue of face presentation attack detection under cross-domain settings (e.g., the setting of unseen attacks and varying illumination) remains to be solved. In this paper, we propose a novel framework based on asymmetric modality translation for face presentation attack detection in bi-modality scenarios. Under the framework, we establish connections between two modality images of genuine faces. Specifically, a novel modality fusion scheme is presented that the image of one modality is translated to the other one through an asymmetric modality translator, then fused with its corresponding paired image. The fusion result is fed as the input to a discriminator for inference. The training of the translator is supervised by an asymmetric modality translation loss. Besides, an illumination normalization module based on Pattern of Local Gravitational Force (PLGF) representation is used to reduce the impact of illumination variation. We conduct extensive experiments on three public datasets, which validate that our method is effective in detecting various types of attacks and achieves state-of-the-art performance under different evaluation protocols.

📄 PDF Abstract BibTeX arXiv:2110.09108

Code (0)

등록된 구현이 없습니다.

Tasks

Face Presentation Attack DetectionFace RecognitionTranslation

Similar Papers 제목 키워드 기반

STARS: Shared-specific Translation and Alignment for missing-modality Remote Sensing Semantic Segmentation

2026-01-24 · Tong Wang, Xiaodong Zhang, Guanzhou Chen, Jiaqi Wang 외 arxiv

Multimodal remote sensing technology significantly enhances the understanding of surface semantics by integrating heterogeneous data such as optical images, Synthetic Aperture Radar (SAR), and Digital Surface Models (DSM…

Semantic Segmentation

AMMNet: An Asymmetric Multi-Modal Network for Remote Sensing Semantic Segmentation

2025-07-22 · Hui Ye, Haodong Chen, Zeke Zexi Hu, Xiaoming Chen 외 arxiv

Semantic segmentation in remote sensing (RS) has advanced significantly with the incorporation of multi-modal data, particularly the integration of RGB imagery and the Digital Surface Model (DSM), which provides compleme…

Semantic Segmentation

MS-MT: Multi-Scale Mean Teacher with Contrastive Unpaired Translation for Cross-Modality Vestibular Schwannoma and Cochlea Segmentation

2023-03-28 · Ziyuan Zhao, Kaixin Xu, Huai Zhe Yeo, Xulei Yang 외

Domain shift has been a long-standing issue for medical image segmentation. Recently, unsupervised domain adaptation (UDA) methods have achieved promising cross-modality segmentation performance by distilling knowledge f…

Domain AdaptationEnsemble LearningImage SegmentationMedical Image Segmentation+3

Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-Spoofing

2024-02-29 · CVPR 2024 1 · Xun Lin, Shuai Wang, Rizhao Cai, Yizhong Liu 외

Face Anti-Spoofing (FAS) is crucial for securing face recognition systems against presentation attacks. With advancements in sensor manufacture and multi-modal learning techniques, many multi-modal FAS approaches have em…

Domain GeneralizationFace Anti-SpoofingFace Recognition

Asymmetric Hierarchical Anchoring for Robust Audio-Visual Cross-Modal Generalization

2026-02-03 · Bixing Wu, Yuhong Zhao, Zongli Ye, Jiachen Lian 외 arxiv

Audio-visual joint representation learning under Cross-Modal Generalization (CMG) aims to transfer knowledge from a labeled source modality to an unlabeled target modality through a unified discrete representation space.…

Representation Learning