paper-with-me

홈 › Papers

Harder or Different? Understanding Generalization of Audio Deepfake Detection

2024-06-05 · Nicolas M. Müller, Nicholas Evans, Hemlata Tak, Philip Sperl, Konstantin Böttinger

Recent research has highlighted a key issue in speech deepfake detection: models trained on one set of deepfakes perform poorly on others. The question arises: is this due to the continuously improving quality of Text-to-Speech (TTS) models, i.e., are newer DeepFakes just 'harder' to detect? Or, is it because deepfakes generated with one model are fundamentally different to those generated using another model? We answer this question by decomposing the performance gap between in-domain and out-of-domain test data into 'hardness' and 'difference' components. Experiments performed using ASVspoof databases indicate that the hardness component is practically negligible, with the performance gap being attributed primarily to the difference component. This has direct implications for real-world deepfake detection, highlighting that merely increasing model capacity, the currently-dominant research trend, may not effectively address the generalization challenge.

📄 PDF Abstract BibTeX arXiv:2406.03512

Code (0)

등록된 구현이 없습니다.

Tasks

Audio Deepfake DetectionDeepFake DetectionFace Swappingtext-to-speechText to Speech

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Does Audio Deepfake Detection Generalize?

2022-03-30 · Nicolas M. Müller, Pavel Czempin, Franziska Dieckmann, Adam Froghyar 외

Current text-to-speech algorithms produce realistic fakes of human voices, making deepfake detection a much-needed area of research. While researchers have presented various techniques for detecting audio spoofs, it is o…

Audio Deepfake DetectionDeepFake DetectionFace Swappingtext-to-speech+1

Human Detection of Political Speech Deepfakes across Transcripts, Audio, and Video

2022-02-25 · Matthew Groh, Aruna Sankaranarayanan, Nikhil Singh, Dong Young Kim 외

Recent advances in technology for hyper-realistic visual and audio effects provoke the concern that deepfake videos of political speeches will soon be indistinguishable from authentic video recordings. The conventional w…

Face SwappingHuman DetectionMisinformationtext-to-speech+1

Audios Don't Lie: Multi-Frequency Channel Attention Mechanism for Audio Deepfake Detection

2024-12-12 · Yangguang Feng

With the rapid development of artificial intelligence technology, the application of deepfake technology in the audio field has gradually increased, resulting in a wide range of security risks. Especially in the financia…

Audio Deepfake DetectionDeepFake DetectionFace Swapping

Attack Agnostic Dataset: Towards Generalization and Stabilization of Audio DeepFake Detection

2022-06-27 · Piotr Kawa, Marcin Plata, Piotr Syga

Audio DeepFakes allow the creation of high-quality, convincing utterances and therefore pose a threat due to its potential applications such as impersonation or fake news. Methods for detecting these manipulations should…

Audio Deepfake DetectionDeepFake DetectionFace Swapping

XMAD-Bench: Cross-Domain Multilingual Audio Deepfake Benchmark

2025-05-31 · Ioan-Paul Ciobanu, Andrei-Iulian Hiji, Nicolae-Catalin Ristea, Paul Irofti 외

Recent advances in audio generation led to an increasing number of deepfakes, making the general public more vulnerable to financial scams, identity theft, and misinformation. Audio deepfake detectors promise to alleviat…

Audio GenerationFace SwappingMisinformation