ADD 2022: the First Audio Deep Synthesis Detection Challenge
Audio deepfake detection is an emerging topic, which was included in the ASVspoof 2021. However, the recent shared tasks have not covered many real-life and challenging scenarios. The first Audio Deep synthesis Detection challenge (ADD) was motivated to fill in the gap. The ADD 2022 includes three tracks: low-quality fake audio detection (LF), partially fake audio detection (PF) and audio fake game (FG). The LF track focuses on dealing with bona fide and fully fake utterances with various real-world noises etc. The PF track aims to distinguish the partially fake audio from the real. The FG track is a rivalry game, which includes two tasks: an audio generation task and an audio fake detection task. In this paper, we describe the datasets, evaluation metrics, and protocols. We also report major findings that reflect the recent advances in audio deepfake detection tasks.
Code (0)
등록된 구현이 없습니다.
Tasks
Audio Deepfake DetectionAudio GenerationDeepFake DetectionFace SwappingSimilar Papers 제목 키워드 기반
Partially Fake Audio Detection by Self-attention-based Fake Span Discovery
The past few years have witnessed the significant advances of speech synthesis and voice conversion technologies. However, such technologies can undermine the robustness of broadly implemented biometric identification mo…
Open-Ended Question AnsweringQuestion AnsweringSpeech SynthesisVoice ConversionWavLM model ensemble for audio deepfake detection
Audio deepfake detection has become a pivotal task over the last couple of years, as many recent speech synthesis and voice cloning systems generate highly realistic speech samples, thus enabling their use in malicious a…
Audio Deepfake DetectionData AugmentationDeepFake DetectionFace Swapping+3Audio Deep Fake Detection System with Neural Stitching for ADD 2022
This paper describes our best system and methodology for ADD 2022: The First Audio Deep Synthesis Detection Challenge\cite{Yi2022ADD}. The very same system was used for both two rounds of evaluation in Track 3.2 with a s…
text-to-speechText to SpeechVoice ConversionDeepfake Detection System for the ADD Challenge Track 3.2 Based on Score Fusion
This paper describes the deepfake audio detection system submitted to the Audio Deep Synthesis Detection (ADD) Challenge Track 3.2 and gives an analysis of score fusion. The proposed system is a score-level fusion of sev…
Data AugmentationDeepFake DetectionFace SwappingThe Vicomtech Audio Deepfake Detection System based on Wav2Vec2 for the 2022 ADD Challenge
This paper describes our submitted systems to the 2022 ADD challenge withing the tracks 1 and 2. Our approach is based on the combination of a pre-trained wav2vec2 feature extractor and a downstream classifier to detect …
Audio Deepfake DetectionAudio SynthesisData AugmentationDeepFake Detection+1