paper-with-me

Papers

SHIELD : An Evaluation Benchmark for Face Spoofing and Forgery Detection with Multimodal Large Language Models

2024-02-06 · Yichen Shi, Yuhao Gao, Yingxin Lai, Hongyang Wang, Jun Feng, Lei He, Jun Wan, Changsheng chen, Zitong Yu, Xiaochun Cao

Multimodal large language models (MLLMs) have demonstrated strong capabilities in vision-related tasks, capitalizing on their visual semantic comprehension and reasoning capabilities. However, their ability to detect subtle visual spoofing and forgery clues in face attack detection tasks remains underexplored. In this paper, we introduce a benchmark, SHIELD, to evaluate MLLMs for face spoofing and forgery detection. Specifically, we design true/false and multiple-choice questions to assess MLLM performance on multimodal face data across two tasks. For the face anti-spoofing task, we evaluate three modalities (i.e., RGB, infrared, and depth) under six attack types. For the face forgery detection task, we evaluate GAN-based and diffusion-based data, incorporating visual and acoustic modalities. We conduct zero-shot and few-shot evaluations in standard and chain of thought (COT) settings. Additionally, we propose a novel multi-attribute chain of thought (MA-COT) paradigm for describing and judging various task-specific and task-irrelevant attributes of face images. The findings of this study demonstrate that MLLMs exhibit strong potential for addressing the challenges associated with the security of facial recognition technology applications.

📄 PDF Abstract BibTeX arXiv:2402.04178

Code (2)

laiyingxin2/shield 공식 구현
FaceOnLive/Face-Liveness-Detection-SDK-Linux

Tasks

AttributeFace Anti-SpoofingMultiple-choiceObject Recognition

Similar Papers 제목 키워드 기반

UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning

2026-05-09 · Hongrui Li, Yichen Shi, Hongyang Wang, Yuhao Gao 외 arxiv

Unified face attack detection (UAD) requires recognizing physical spoofing and digital forgery within a shared decision space, yet existing discriminative or prompt-based methods largely rely on appearance correlations a…

Multimodal Reasoning

FaceShield: Explainable Face Anti-Spoofing with Multimodal Large Language Models

2025-05-14 · Hongyang Wang, Yichen Shi, Zhuofu Tao, Yuhao Gao 외

Face anti-spoofing (FAS) is crucial for protecting facial recognition systems from presentation attacks. Previous methods approached this task as a classification problem, lacking interpretability and reasoning behind th…

Face Anti-Spoofing

Benchmarking Joint Face Spoofing and Forgery Detection with Visual and Physiological Cues

2022-08-10 · Zitong Yu, Rizhao Cai, Zhi Li, Wenhan Yang 외

Face anti-spoofing (FAS) and face forgery detection play vital roles in securing face biometric systems from presentation attacks (PAs) and vicious digital manipulation (e.g., deepfakes). Despite promising performance up…

BenchmarkingDeepFake DetectionFace Anti-SpoofingFace Swapping+1

DeepShield: Fortifying Deepfake Video Detection with Local and Global Forgery Analysis

2025-10-29 · Yinqi Cai, Jichang Li, Zhaolun Li, Weikai Chen 외 arxiv

Recent advances in deep generative models have made it easier to manipulate face videos, raising significant concerns about their potential misuse for fraud and misinformation. Existing detectors often perform well in in…

DeepFake Detection

FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models

2024-10-03 · Zhipei Xu, Xuanyu Zhang, Runyi Li, Zecheng Tang 외

The rapid development of generative AI is a double-edged sword, which not only facilitates content creation but also makes image manipulation easier and more difficult to detect. Although current image forgery detection …

Face SwappingImage Forgery DetectionImage ManipulationTAG