paper-with-me

홈 › Papers

One-Class Learning with Adaptive Centroid Shift for Audio Deepfake Detection

2024-06-24 · Hyun Myung Kim, Kangwook Jang, Hoirin Kim

As speech synthesis systems continue to make remarkable advances in recent years, the importance of robust deepfake detection systems that perform well in unseen systems has grown. In this paper, we propose a novel adaptive centroid shift (ACS) method that updates the centroid representation by continually shifting as the weighted average of bonafide representations. Our approach uses only bonafide samples to define their centroid, which can yield a specialized centroid for one-class learning. Integrating our ACS with one-class learning gathers bonafide representations into a single cluster, forming well-separated embeddings robust to unseen spoofing attacks. Our proposed method achieves an equal error rate (EER) of 2.19% on the ASVspoof 2021 deepfake dataset, outperforming all existing systems. Furthermore, the t-SNE visualization illustrates that our method effectively maps the bonafide embeddings into a single cluster and successfully disentangles the bonafide and spoof classes.

📄 PDF Abstract BibTeX arXiv:2406.16716

Code (1)

avishai111/One-Class-Learning-with-Adaptive-Centroid-Shift-for-Audio-Deepfake-Detection pytorch

Tasks

Audio Deepfake DetectionDeepFake DetectionFace SwappingSpeech Synthesis

Similar Papers 제목 키워드 기반

From Talking to Singing: A New Challenge for Audio-Visual Deepfake Detection

2026-05-27 · Ke Liu, Jiwei Wei, Wenyu Zhang, Shuchang Zhou 외 arxiv

With rapid advances in audio-visual generative models, reliable forgery detection becomes increasingly critical. Existing methods for audio-visual deepfake detection typically rely on cross-modal inconsistencies. In sing…

DeepFake Detection

QAMO: Quality-aware Multi-centroid One-class Learning For Speech Deepfake Detection

2025-09-25 · Duc-Tuan Truong, Tianchi Liu, Ruijie Tao, Junjie Li 외 arxiv

Recent work shows that one-class learning can detect unseen deepfake attacks by modeling a compact distribution of bona fide speech around a single centroid. However, the single-centroid assumption can oversimplify the b…

DeepFake Detection

What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection

2023-12-15 · Xiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chuyuan Zhang 외

The rapid evolution of speech synthesis and voice conversion has raised substantial concerns due to the potential misuse of such technology, prompting a pressing need for effective audio deepfake detection mechanisms. Ex…

Audio Deepfake DetectionContinual LearningDeepFake DetectionFace Swapping+2

Adversarial Attacks on Audio Deepfake Detection: A Benchmark and Comparative Study

2025-09-08 · Kutub Uddin, Muhammad Umar Farooq, Awais Khan, Khalid Mahmood Malik arxiv

The widespread use of generative AI has shown remarkable success in producing highly realistic deepfakes, posing a serious threat to various voice biometric applications, including speaker verification, voice biometrics,…

Audio Deepfake DetectionSpeaker Verification

Reject Threshold Adaptation for Open-Set Model Attribution of Deepfake Audio

2024-12-02 · Xinrui Yan, Jiangyan Yi, JianHua Tao, Yujie Chen 외

Open environment oriented open set model attribution of deepfake audio is an emerging research topic, aiming to identify the generation models of deepfake audio. Most previous work requires manually setting a rejection t…

Face Swapping