paper-with-me

Papers

DeepFake Doctor: Diagnosing and Treating Audio-Video Fake Detection

2025-06-06 · Marcel Klemt, Carlotta Segna, Anna Rohrbach

Generative AI advances rapidly, allowing the creation of very realistic manipulated video and audio. This progress presents a significant security and ethical threat, as malicious users can exploit DeepFake techniques to spread misinformation. Recent DeepFake detection approaches explore the multimodal (audio-video) threat scenario. In particular, there is a lack of reproducibility and critical issues with existing datasets - such as the recently uncovered silence shortcut in the widely used FakeAVCeleb dataset. Considering the importance of this topic, we aim to gain a deeper understanding of the key issues affecting benchmarking in audio-video DeepFake detection. We examine these challenges through the lens of the three core benchmarking pillars: datasets, detection methods, and evaluation protocols. To address these issues, we spotlight the recent DeepSpeak v1 dataset and are the first to propose an evaluation protocol and benchmark it using SOTA models. We introduce SImple Multimodal BAseline (SIMBA), a competitive yet minimalistic approach that enables the exploration of diverse design choices. We also deepen insights into the issue of audio shortcuts and present a promising mitigation strategy. Finally, we analyze and enhance the evaluation scheme on the widely used FakeAVCeleb dataset. Our findings offer a way forward in the complex area of audio-video DeepFake detection.

📄 PDF Abstract BibTeX arXiv:2506.05851

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingDeepFake DetectionFace SwappingMisinformation

Similar Papers 제목 키워드 기반

Model Doctor: A Simple Gradient Aggregation Strategy for Diagnosing and Treating CNN Classifiers

2021-12-09 · Zunlei Feng, Jiacong Hu, Sai Wu, Xiaotian Yu 외

Recently, Convolutional Neural Network (CNN) has achieved excellent performance in the classification task. It is widely known that CNN is deemed as a 'black-box', which is hard for understanding the prediction mechanism…

Prediction

FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

2021-08-11 · Hasam Khalid, Shahroz Tariq, Minha Kim, Simon S. Woo

While the significant advancements have made in the generation of deepfakes using deep learning technologies, its misuse is a well-known issue now. Deepfakes can cause severe security and privacy issues as they can be us…

DeepFake DetectionFace Swapping

Evaluation of an Audio-Video Multimodal Deepfake Dataset using Unimodal and Multimodal Detectors

2021-09-07 · Hasam Khalid, Minha Kim, Shahroz Tariq, Simon S. Woo

Significant advancements made in the generation of deepfakes have caused security and privacy issues. Attackers can easily impersonate a person's identity in an image by replacing his face with the target person's face. …

DeepFake DetectionFace Swapping

Model Doctor for Diagnosing and Treating Segmentation Error

2023-02-17 · Zhijie Jia, Lin Chen, Kaiwen Hu, Lechao Cheng 외

Despite the remarkable progress in semantic segmentation tasks with the advancement of deep neural networks, existing U-shaped hierarchical typical segmentation networks still suffer from local misclassification of categ…

modelSegmentationSemantic Segmentation

How Deep Are the Fakes? Focusing on Audio Deepfake: A Survey

2021-11-28 · Zahra Khanjani, Gabrielle Watson, Vandana P. Janeja

Deepfake is content or material that is synthetically generated or manipulated using artificial intelligence (AI) methods, to be passed off as real and can include audio, video, image, and text synthesis. This survey has…

Face SwappingFake News DetectionSurvey