paper-with-me

Papers

Integrating Audio-Visual Features for Multimodal Deepfake Detection

2023-10-05 · Sneha Muppalla, Shan Jia, Siwei Lyu

Deepfakes are AI-generated media in which an image or video has been digitally modified. The advancements made in deepfake technology have led to privacy and security issues. Most deepfake detection techniques rely on the detection of a single modality. Existing methods for audio-visual detection do not always surpass that of the analysis based on single modalities. Therefore, this paper proposes an audio-visual-based method for deepfake detection, which integrates fine-grained deepfake identification with binary classification. We categorize the samples into four types by combining labels specific to each single modality. This method enhances the detection under intra-domain and cross-domain testing.

📄 PDF Abstract BibTeX arXiv:2310.03827

Code (0)

등록된 구현이 없습니다.

Tasks

Binary ClassificationDeepFake DetectionFace Swapping

Similar Papers 제목 키워드 기반

MIS-AVoiDD: Modality Invariant and Specific Representation for Audio-Visual Deepfake Detection

2023-10-03 · Vinaya Sree Katamneni, Ajita Rattani

Deepfakes are synthetic media generated using deep generative algorithms and have posed a severe societal and political threat. Apart from facial manipulation and synthetic voice, recently, a novel kind of deepfakes has …

DeepFake DetectionFace Swapping

KLASSify to Verify: Audio-Visual Deepfake Detection Using SSL-based Audio and Handcrafted Visual Features

2025-08-10 · Ivan Kukanov, Jun Wah Ng arxiv

The rapid development of audio-driven talking head generators and advanced Text-To-Speech (TTS) models has led to more sophisticated temporal deepfakes. These advances highlight the need for robust methods capable of det…

Self-Supervised LearningDeepFake Detection

AV-Lip-Sync+: Leveraging AV-HuBERT to Exploit Multimodal Inconsistency for Video Deepfake Detection

2023-11-05 · Sahibzada Adil Shahzad, Ammarah Hashmi, Yan-Tsung Peng, Yu Tsao 외

Multimodal manipulations (also known as audio-visual deepfakes) make it difficult for unimodal deepfake detectors to detect forgeries in multimedia content. To avoid the spread of false propaganda and fake news, timely d…

DeepFake DetectionFace SwappingSelf-Supervised LearningVideo Forensics

DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts

2026-08-24 · Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu, A. Sophia Koepke 외 arxiv

Audio-visual deepfake detection is an actively studied topic, where one of the main challenges is to develop detectors able to generalize across deepfake generation methods. We conjecture that overfitting can be mitigate…

DeepFake DetectionFace Parsing

ERF-BA-TFD+: A Multimodal Model for Audio-Visual Deepfake Detection

2025-08-24 · Xin Zhang, Jiaming Chu, Jian Zhao, Yuchu Jiang 외 arxiv

Deepfake detection is a critical task in identifying manipulated multimedia content. In real-world scenarios, deepfake content can manifest across multiple modalities, including audio and video. To address this challenge…

DeepFake Detection