Advancing Continual Learning for Robust Deepfake Audio Classification
The emergence of new spoofing attacks poses an increasing challenge to audio security. Current detection methods often falter when faced with unseen spoofing attacks. Traditional strategies, such as retraining with new data, are not always feasible due to extensive storage. This paper introduces a novel continual learning method Continual Audio Defense Enhancer (CADE). First, by utilizing a fixed memory size to store randomly selected samples from previous datasets, our approach conserves resources and adheres to privacy constraints. Additionally, we also apply two distillation losses in CADE. By distillation in classifiers, CADE ensures that the student model closely resembles that of the teacher model. This resemblance helps the model retain old information while facing unseen data. We further refine our model's performance with a novel embedding similarity loss that extends across multiple depth layers, facilitating superior positive sample alignment. Experiments conducted on the ASVspoof2019 dataset show that our proposed method outperforms the baseline methods.
Code (0)
등록된 구현이 없습니다.
Tasks
Audio ClassificationClassificationContinual LearningFace SwappingSimilar Papers 제목 키워드 기반
What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection
The rapid evolution of speech synthesis and voice conversion has raised substantial concerns due to the potential misuse of such technology, prompting a pressing need for effective audio deepfake detection mechanisms. Ex…
Audio Deepfake DetectionContinual LearningDeepFake DetectionFace Swapping+2Region-Based Optimization in Continual Learning for Audio Deepfake Detection
Rapid advancements in speech synthesis and voice conversion bring convenience but also new security risks, creating an urgent need for effective audio deepfake detection. Although current models perform well, their effec…
Audio Deepfake DetectionContinual LearningDeepFake DetectionFace Swapping+2Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
Currently, Audio Language Models (ALMs) are rapidly advancing due to the developments in large language models and audio neural codecs. These ALMs have significantly lowered the barrier to creating deepfake audio, genera…
Audio Deepfake DetectionDeepFake DetectionFace SwappingBanglaFake: Constructing and Evaluating a Specialized Bengali Deepfake Audio Dataset
Deepfake audio detection is challenging for low-resource languages like Bengali due to limited datasets and subtle acoustic features. To address this, we introduce BangalFake, a Bengali Deepfake Audio Dataset with 12,260…
DeepFake DetectionFace Swappingtext-to-speechText to SpeechRehearsal with Auxiliary-Informed Sampling for Audio Deepfake Detection
The performance of existing audio deepfake detection frameworks degrades when confronted with new deepfake attacks. Rehearsal-based continual learning (CL), which updates models using a limited set of old data samples, h…
Audio Deepfake DetectionContinual LearningDeepFake DetectionDiversity+1