paper-with-me

홈 › Papers

Lip Sync Matters: A Novel Multimodal Forgery Detector

2022-11-07 · APSIPA ASC 2022 2022 11 · Sahibzada Adil Shahzad, Ammarah Hashmi, Sarwar Khan, Yan-Tsung Peng, Yu Tsao, Hsin-Min Wang

Deepfake technology has advanced a lot, but it is a double-sided sword for the community. One can use it for beneficial purposes, such as restoring vintage content in old movies, or for nefarious purposes, such as creating fake footage to manipulate the public and distribute non-consensual pornography. A lot of work has been done to combat its improper use by detecting fake footage with good performance thanks to the availability of numerous public datasets and unimodal deep learning-based models. However, these methods are insufficient to detect multimodal manipulations, such as both visual and acoustic. This work proposes a novel lip-reading-based multi-modal Deepfake detection method called “Lip Sync Matters.” It targets high-level semantic features to exploit the mismatch between the lip sequence extracted from the video and the synthetic lip sequence generated from the audio by the Wav2lip model to detect forged videos. Experimental results show that the proposed method outperforms several existing unimodal, ensemble, and multimodal methods on the publicly available multimodal FakeAVCeleb dataset.

📄 PDF Abstract BibTeX

Code (1)

sahibzadaadil/Lip-Sync-Matters-A-Novel-Multimodal-Forgery-Detector pytorch

Tasks

DeepFake DetectionFace SwappingLip Reading

Similar Papers 제목 키워드 기반

Frequency Bias Matters: Diving into Robust and Generalized Deep Image Forgery Detection

2025-11-25 · Chi Liu, Tianqing Zhu, Wanlei Zhou, Wei Zhao arxiv

As deep image forgery powered by AI generative models, such as GANs, continues to challenge today's digital world, detecting AI-generated forgeries has become a vital security topic. Generalizability and robustness are t…

Forensics-Bench: A Comprehensive Forgery Detection Benchmark Suite for Large Vision Language Models

2025-03-19 · CVPR 2025 1 · Jin Wang, Chenghui Lv, Xian Li, Shichao Dong 외

Recently, the rapid development of AIGC has significantly boosted the diversities of fake media spread in the Internet, posing unprecedented threats to social security, politics, law, and etc. To detect the ever-increasi…

AV-Lip-Sync+: Leveraging AV-HuBERT to Exploit Multimodal Inconsistency for Video Deepfake Detection

2023-11-05 · Sahibzada Adil Shahzad, Ammarah Hashmi, Yan-Tsung Peng, Yu Tsao 외

Multimodal manipulations (also known as audio-visual deepfakes) make it difficult for unimodal deepfake detectors to detect forgeries in multimedia content. To avoid the spread of false propaganda and fake news, timely d…

DeepFake DetectionFace SwappingSelf-Supervised LearningVideo Forensics

FKA-Owl: Advancing Multimodal Fake News Detection through Knowledge-Augmented LVLMs

2024-03-04 · Xuannan Liu, Peipei Li, Huaibo Huang, Zekun Li 외

The massive generation of multimodal fake news involving both text and images exhibits substantial distribution discrepancies, prompting the need for generalized detectors. However, the insulated nature of training restr…

Fake News DetectionImage ManipulationInformativenessWorld Knowledge

DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection

2025-10-26 · Kangran Zhao, Yupeng Chen, Xiaoyu Zhang, Yize Chen 외 arxiv

The misuse of advanced generative AI models has resulted in the widespread proliferation of falsified data, particularly forged human-centric audiovisual content, which poses substantial societal risks (e.g., financial f…

DeepFake Detection