Synthetic speech detection using meta-learning with prototypical loss
Recent works on speech spoofing countermeasures still lack generalization ability to unseen spoofing attacks. This is one of the key issues of ASVspoof challenges especially with the rapid development of diverse and high-quality spoofing algorithms. In this work, we address the generalizability of spoofing detection by proposing prototypical loss under the meta-learning paradigm to mimic the unseen test scenario during training. Prototypical loss with metric-learning objectives can learn the embedding space directly and emerges as a strong alternative to prevailing classification loss functions. We propose an anti-spoofing system based on squeeze-excitation Residual network (SE-ResNet) architecture with prototypical loss. We demonstrate that the proposed single system without any data augmentation can achieve competitive performance to the recent best anti-spoofing systems on ASVspoof 2019 logical access (LA) task. Furthermore, the proposed system with data augmentation outperforms the ASVspoof 2021 challenge best baseline both in the progress and evaluation phase of the LA task. On ASVspoof 2019 and 2021 evaluation set LA scenario, we attain a relative 68.4% and 3.6% improvement in min-tDCF compared to the challenge best baselines, respectively.
Code (0)
등록된 구현이 없습니다.
Tasks
Data AugmentationMeta-LearningMetric LearningSynthetic Speech DetectionSimilar Papers 제목 키워드 기반
Anomaly Detection in Human Language via Meta-Learning: A Few-Shot Approach
We propose a meta learning framework for detecting anomalies in human language across diverse domains with limited labeled data. Anomalies in language ranging from spam and fake news to hate speech pose a major challenge…
Binary ClassificationAnomaly DetectionProtAugment: Intent Detection Meta-Learning through Unsupervised Diverse Paraphrasing
Recent research considers few-shot intent detection as a meta-learning problem: the model is learning to learn from a consecutive set of small tasks named episodes. In this work, we propose ProtAugment, a meta-learning a…
DiversityIntent DetectionLanguage ModelingLanguage Modelling+1ProtAugment: Unsupervised diverse short-texts paraphrasing for intent detection meta-learning
Recent research considers few-shot intent detection as a meta-learning problem: the model is learning to learn from a consecutive set of small tasks named episodes. In this work, we propose ProtAugment, a meta-learning a…
DiversityIntent DetectionLanguage ModelingLanguage Modelling+1Improved Meta-Learning Training for Speaker Verification
Meta-learning has recently become a research hotspot in speaker verification (SV). We introduce two methods to improve the meta-learning training for SV in this paper. For the first method, a backbone embedding network i…
Data AugmentationMeta-LearningSpeaker Verificationspeech-recognition+1"Define Your Terms" : Enhancing Efficient Offensive Speech Classification with Definition
The propagation of offensive content through social media channels has garnered attention of the research community. Multiple works have proposed various semantically related yet subtle distinct categories of offensive s…
Diversity