paper-with-me

Papers

Multiscale Adaptive Conflict-Balancing Model For Multimedia Deepfake Detection

2025-05-19 · Zihan Xiong, Xiaohua Wu, Lei Chen, Fangqi Lou

Advances in computer vision and deep learning have blurred the line between deepfakes and authentic media, undermining multimedia credibility through audio-visual forgery. Current multimodal detection methods remain limited by unbalanced learning between modalities. To tackle this issue, we propose an Audio-Visual Joint Learning Method (MACB-DF) to better mitigate modality conflicts and neglect by leveraging contrastive learning to assist in multi-level and cross-modal fusion, thereby fully balancing and exploiting information from each modality. Additionally, we designed an orthogonalization-multimodal pareto module that preserves unimodal information while addressing gradient conflicts in audio-video encoders caused by differing optimization targets of the loss functions. Extensive experiments and ablation studies conducted on mainstream deepfake datasets demonstrate consistent performance gains of our model across key evaluation metrics, achieving an average accuracy of 95.5% across multiple datasets. Notably, our method exhibits superior cross-dataset generalization capabilities, with absolute improvements of 8.0% and 7.7% in ACC scores over the previous best-performing approach when trained on DFDC and tested on DefakeAVMiT and FakeAVCeleb datasets.

📄 PDF Abstract BibTeX arXiv:2505.12966

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningDeepFake DetectionFace Swapping

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Inclusion 2024 Global Multimedia Deepfake Detection Challenge: Towards Multi-dimensional Face Forgery Detection

2024-12-30 · Yi Zhang, Weize Gao, Changtao Miao, Man Luo 외

In this paper, we present the Global Multimedia Deepfake Detection held concurrently with the Inclusion 2024. Our Multimedia Deepfake Detection aims to detect automatic image and audio-video manipulations including but n…

DeepFake DetectionFace Swappingvalid

Grand Challenge On Detecting Cheapfakes

2023-04-03 · Duc-Tien Dang-Nguyen, Sohail Ahmed Khan, Cise Midoglu, Michael Riegler 외

Cheapfake is a recently coined term that encompasses non-AI ("cheap") manipulations of multimedia content. Cheapfakes are known to be more prevalent than deepfakes. Cheapfake media can be created using editing software f…

Image Captioning

Suppressing Gradient Conflict for Generalizable Deepfake Detection

2025-07-29 · Ming-Hui Liu, Harry Cheng, Xin Luo, Xin-Shun Xu arxiv

Robust deepfake detection models must be capable of generalizing to ever-evolving manipulation techniques beyond training data. A promising strategy is to augment the training data with online synthesized fake images con…

Representation LearningDomain GeneralizationDeepFake Detection

Balancing multiscale similarity and cartographic constraints: A similarity-driven optimization framework for line generalization

2026-07-28 · Pengbo Li, Haowen Yan, Xiaomin Lu, Binbin Lin arxiv

Cartographic generalization is essential for generating multiscale map representations by balancing information preservation and cartographic readability. However, automated generalization remains challenging because exi…

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan

2026-04-09 · Yuankun Xie, Haonan Cheng, Jiayi Zhou, Xiaoxuan Guo 외 arxiv

The rapid advancement of Audio Large Language Models (ALLMs) has enabled cost-effective, high-fidelity generation and manipulation of both speech and non-speech audio, including sound effects, singing voices, and music. …

Audio Deepfake Detection