paper-with-me

홈 › Papers

CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-spoofing

2024-03-21 · CVPR 2024 1 · Ajian Liu, Shuai Xue, Jianwen Gan, Jun Wan, Yanyan Liang, Jiankang Deng, Sergio Escalera, Zhen Lei

Domain generalization (DG) based Face Anti-Spoofing (FAS) aims to improve the model's performance on unseen domains. Existing methods either rely on domain labels to align domain-invariant feature spaces, or disentangle generalizable features from the whole sample, which inevitably lead to the distortion of semantic feature structures and achieve limited generalization. In this work, we make use of large-scale VLMs like CLIP and leverage the textual feature to dynamically adjust the classifier's weights for exploring generalizable visual features. Specifically, we propose a novel Class Free Prompt Learning (CFPL) paradigm for DG FAS, which utilizes two lightweight transformers, namely Content Q-Former (CQF) and Style Q-Former (SQF), to learn the different semantic prompts conditioned on content and style features by using a set of learnable query vectors, respectively. Thus, the generalizable prompt can be learned by two improvements: (1) A Prompt-Text Matched (PTM) supervision is introduced to ensure CQF learns visual representation that is most informative of the content description. (2) A Diversified Style Prompt (DSP) technology is proposed to diversify the learning of style prompts by mixing feature statistics between instance-specific styles. Finally, the learned text features modulate visual features to generalization through the designed Prompt Modulation (PM). Extensive experiments show that the CFPL is effective and outperforms the state-of-the-art methods on several cross-domain datasets.

📄 PDF Abstract BibTeX arXiv:2403.14333

Code (0)

등록된 구현이 없습니다.

Tasks

Domain GeneralizationFace Anti-SpoofingPrompt Learning

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

BCFPL: Binary classification ConvNet based Fast Parking space recognition with Low resolution image

2024-04-22 · Shuo Zhang, Xin Chen, Zixuan Wang

The automobile plays an important role in the economic activities of mankind, especially in the metropolis. Under the circumstances, the demand of quick search for available parking spaces has become a major concern for …

Binary Classification

Efficient Brood Cell Detection in Layer Trap Nests for Bees and Wasps: Balancing Labeling Effort and Species Coverage

2026-03-17 · Chenchang Liu, Felix Fornoff, Annika Grasreiner, Patrick Maeder 외 arxiv

Monitoring cavity-nesting wild bees and wasps is vital for biodiversity research and conservation. Layer trap nests (LTNs) are emerging as a valuable tool to study the abundance and species richness of these insects, off…

Cell Detection

PPOM: Marginalizing Patch-Grid Phase for CLIP-Based Generalizable Vision-Language Prompt Tuning

2026-08-14 · Liang Wang, Haoyang Li, Chao Wang, Guodong Long 외 arxiv

Prompt tuning adapts CLIP-based vision-language models with few trainable parameters, yet its predictions remain sensitive to the spatial sampling imposed by a frozen vision transformer. In particular, non-overlapping pa…

MADPromptS: Unlocking Zero-Shot Morphing Attack Detection with Multiple Prompt Aggregation

2025-08-12 · Eduarda Caldeira, Fadi Boutros, Naser Damer arxiv

Face Morphing Attack Detection (MAD) is a critical challenge in face recognition security, where attackers can fool systems by interpolating the identity information of two or more individuals into a single face image, r…

Prompt EngineeringFace Recognition

PanSAM: Zero-Shot, Prompt-Free Pancreas Segmentation in CT Imaging

2024-07-03 · ICML 2024 FM-Wild Workshop 2024 7 · Abolfazl malekahmadi, Mohammad Taha Teimuri Jervakani, Armin Behnamnia, Zahra Dehghanian 외

Segmentation of the pancreas in CT images is crucial in multiple pancreatic diagnostic tasks, such as the detection, classification, and prognosis of pancreatic cancer. We present a segmentation model to find pancreatic …

DiagnosticPancreas SegmentationPrognosisSegmentation+1