Towards More General Video-based Deepfake Detection through Facial Component Guided Adaptation for Foundation Model
The current deep generative models have enabled the creation of synthetic facial images with remarkable photorealism, raising significant societal concerns over their potential misuse. Despite rapid advancements in the field of deepfake detection, developing an efficient and effective approach for the generalized deepfake detection of unseen forgery samples remains challenging. To address this challenge, we leverage the rich semantic priors of foundation models and propose a novel side-network-based decoder that extracts spatial and temporal cues using the CLIP image encoder for generalized video-based Deepfake detection. Additionally, we introduce Facial Component Guidance (FCG) to enhance spatial learning generalizability by encouraging the model to focus on key facial regions. By leveraging the generic features of a vision-language foundation model, our approach demonstrates promising generalizability on challenging Deepfake datasets while also exhibiting superiority in training data efficiency, parameter efficiency, and model robustness. The source code is available at: https://github.com/aiiu-lab/DFD-FCG.
Code (1)
Tasks
DecoderDeepFake DetectionFace SwappingMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
One Detector to Rule Them All: Towards a General Deepfake Attack Detection Framework
Deep learning-based video manipulation methods have become widely accessible to the masses. With little to no effort, people can quickly learn how to generate deepfake (DF) videos. While deep learning-based detection met…
AllDeep LearningFace SwappingDetecting Deepfake by Creating Spatio-Temporal Regularity Disruption
Despite encouraging progress in deepfake detection, generalization to unseen forgery types remains a significant challenge due to the limited forgery clues explored during training. In contrast, we notice a common phenom…
DeepFake DetectionFace SwappingA Convolutional LSTM based Residual Network for Deepfake Video Detection
In recent years, deep learning-based video manipulation methods have become widely accessible to masses. With little to no effort, people can easily learn how to generate deepfake videos with only a few victims or target…
DeepFake DetectionFace SwappingTransfer LearningTowards More General Video-based Deepfake Detection through Facial Feature Guided Adaptation for Foundation Model
With the rise of deep learning, generative models have enabled the creation of highly realistic synthetic images, presenting challenges due to their potential misuse. While research in Deepfake detection has grown rapidl…
DecoderDeepFake DetectionFace Swappingparameter-efficient fine-tuningJoint Audio-Visual Attention with Contrastive Learning for More General Deepfake Detection
With the continuous advancement of deepfake technology, there has been a surge in the creation of realistic fake videos. Unfortunately, the malicious utilization of deepfake poses a significant threat to societal moralit…
Contrastive LearningDeepFake DetectionFace SwappingHuman Detection of Deepfakes