paper-with-me

Papers

GM-DF: Generalized Multi-Scenario Deepfake Detection

2024-06-28 · Yingxin Lai, Zitong Yu, Jing Yang, Bin Li, Xiangui Kang, Linlin Shen

Existing face forgery detection usually follows the paradigm of training models in a single domain, which leads to limited generalization capacity when unseen scenarios and unknown attacks occur. In this paper, we elaborately investigate the generalization capacity of deepfake detection models when jointly trained on multiple face forgery detection datasets. We first find a rapid degradation of detection accuracy when models are directly trained on combined datasets due to the discrepancy across collection scenarios and generation methods. To address the above issue, a Generalized Multi-Scenario Deepfake Detection framework (GM-DF) is proposed to serve multiple real-world scenarios by a unified model. First, we propose a hybrid expert modeling approach for domain-specific real/forgery feature extraction. Besides, as for the commonality representation, we use CLIP to extract the common features for better aligning visual and textual features across domains. Meanwhile, we introduce a masked image reconstruction mechanism to force models to capture rich forged details. Finally, we supervise the models via a domain-aware meta-learning strategy to further enhance their generalization capacities. Specifically, we design a novel domain alignment loss to strongly align the distributions of the meta-test domains and meta-train domains. Thus, the updated models are able to represent both specific and common real/forgery features across multiple datasets. In consideration of the lack of study of multi-dataset training, we establish a new benchmark leveraging multi-source data to fairly evaluate the models' generalization capacity on unseen scenarios. Both qualitative and quantitative experiments on five datasets conducted on traditional protocols as well as the proposed benchmark demonstrate the effectiveness of our approach.

📄 PDF Abstract BibTeX arXiv:2406.20078

Code (1)

laiyingxin2/gm-df 공식 구현

Tasks

DeepFake DetectionFace SwappingImage ReconstructionMeta-Learning

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

TAR: Generalized Forensic Framework to Detect Deepfakes using Weakly Supervised Learning

2021-05-13 · Sangyup Lee, Shahroz Tariq, Junyaup Kim, Simon S. Woo

Deepfakes have become a critical social problem, and detecting them is of utmost importance. Also, deepfake generation methods are advancing, and it is becoming harder to detect. While many deepfake detection models can …

DeepFake DetectionFace SwappingTARTransfer Learning+1

EEG-Features for Generalized Deepfake Detection

2024-05-14 · Arian Beckmann, Tilman Stephani, Felix Klotzsche, Yonghao Chen 외

Since the advent of Deepfakes in digital media, the development of robust and reliable detection mechanism is urgently called for. In this study, we explore a novel approach to Deepfake detection by utilizing electroence…

DeepFake DetectionEEGFace Swapping

Towards Open-world Generalized Deepfake Detection: General Feature Extraction via Unsupervised Domain Adaptation

2025-05-18 · Midou Guo, Qilin Yin, Wei Lu, Xiangyang Luo

With the development of generative artificial intelligence, new forgery methods are rapidly emerging. Social platforms are flooded with vast amounts of unlabeled synthetic data and authentic data, making it increasingly …

DeepFake DetectionDomain AdaptationFace SwappingUnsupervised Domain Adaptation

The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio

2024-05-08 · Yuankun Xie, Yi Lu, Ruibo Fu, Zhengqi Wen 외

With the proliferation of Audio Language Model (ALM) based deepfake audio, there is an urgent need for generalized detection methods. ALM-based deepfake audio currently exhibits widespread, high deception, and type versa…

Audio Deepfake DetectionAudio GenerationDeepFake DetectionFace Swapping+2

Towards More General Video-based Deepfake Detection through Facial Component Guided Adaptation for Foundation Model

2025-01-01 · CVPR 2025 1 · Yue-Hua Han, Tai-Ming Huang, Kai-Lung Hua, Jun-Cheng Chen

The current deep generative models have enabled the creation of synthetic facial images with remarkable photorealism, raising significant societal concerns over their potential misuse. Despite rapid advancements in t…

DecoderDeepFake DetectionFace Swapping