paper-with-me

Papers

Explicit Correlation Learning for Generalizable Cross-Modal Deepfake Detection

2024-04-30 · Cai Yu, Shan Jia, Xiaomeng Fu, Jin Liu, Jiahe Tian, Jiao Dai, Xi Wang, Siwei Lyu, Jizhong Han

With the rising prevalence of deepfakes, there is a growing interest in developing generalizable detection methods for various types of deepfakes. While effective in their specific modalities, traditional detection methods fall short in addressing the generalizability of detection across diverse cross-modal deepfakes. This paper aims to explicitly learn potential cross-modal correlation to enhance deepfake detection towards various generation scenarios. Our approach introduces a correlation distillation task, which models the inherent cross-modal correlation based on content information. This strategy helps to prevent the model from overfitting merely to audio-visual synchronization. Additionally, we present the Cross-Modal Deepfake Dataset (CMDFD), a comprehensive dataset with four generation methods to evaluate the detection of diverse cross-modal deepfakes. The experimental results on CMDFD and FakeAVCeleb datasets demonstrate the superior generalizability of our method over existing state-of-the-art methods. Our code and data can be found at \url{https://github.com/ljj898/CMDFD-Dataset-and-Deepfake-Detection}.

📄 PDF Abstract BibTeX arXiv:2404.19171

Code (1)

ljj898/cmdfd-dataset-and-deepfake-detection 공식 구현 pytorch

Tasks

Audio-Visual SynchronizationDeepFake DetectionFace Swapping

Similar Papers 제목 키워드 기반

Towards Generalizable Deepfake Detection via Forgery-aware Audio-Visual Adaptation: A Variational Bayesian Approach

2025-11-24 · Fan Nie, Jiangqun Ni, Jian Zhang, Bin Zhang 외 arxiv

The widespread application of AIGC contents has brought not only unprecedented opportunities, but also potential security concerns, e.g., audio-visual deepfakes. Therefore, it is of great importance to develop an effecti…

DeepFake Detection

AVFF: Audio-Visual Feature Fusion for Video Deepfake Detection

2024-06-05 · CVPR 2024 1 · Trevine Oorloff, Surya Koppisetti, Nicolò Bonettini, Divyaraj Solanki 외

With the rapid growth in deepfake video content, we require improved and generalizable methods to detect them. Most existing detection methods either use uni-modal cues or rely on supervised training to capture the disso…

Contrastive LearningDeepFake DetectionFace SwappingRepresentation Learning

DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts

2026-08-24 · Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu, A. Sophia Koepke 외 arxiv

Audio-visual deepfake detection is an actively studied topic, where one of the main challenges is to develop detectors able to generalize across deepfake generation methods. We conjecture that overfitting can be mitigate…

DeepFake DetectionFace Parsing

Unsupervised Multimodal Deepfake Detection Using Intra- and Cross-Modal Inconsistencies

2023-11-28 · Mulin Tian, Mahyar Khayatkhoei, Joe Mathai, Wael AbdAlmageed

Deepfake videos present an increasing threat to society with potentially negative impact on criminal justice, democracy, and personal safety and privacy. Meanwhile, detecting deepfakes, at scale, remains a very challengi…

DeepFake DetectionFace Swapping

Leave No Stone Unturned: Uncovering Holistic Audio-Visual Intrinsic Coherence for Deepfake Detection

2026-03-25 · Jielun Peng, Yabin Wang, Yaqi Li, Long Kong 외 arxiv

The rapid progress of generative AI has enabled hyper-realistic audio-visual deepfakes, intensifying threats to personal security and social trust. Most existing deepfake detectors rely either on uni-modal artifacts or a…

DeepFake Detection