paper-with-me

홈 › Papers

Audio Spoofing Verification using Deep Convolutional Neural Networks by Transfer Learning

2020-08-08 · Rahul T P, P R Aravind, Ranjith C, Usamath Nechiyil, Nandakumar Paramparambath

Automatic Speaker Verification systems are gaining popularity these days; spoofing attacks are of prime concern as they make these systems vulnerable. Some spoofing attacks like Replay attacks are easier to implement but are very hard to detect thus creating the need for suitable countermeasures. In this paper, we propose a speech classifier based on deep-convolutional neural network to detect spoofing attacks. Our proposed methodology uses acoustic time-frequency representation of power spectral densities on Mel frequency scale (Mel-spectrogram), via deep residual learning (an adaptation of ResNet-34 architecture). Using a single model system, we have achieved an equal error rate (EER) of 0.9056% on the development and 5.32% on the evaluation dataset of logical access scenario and an equal error rate (EER) of 5.87% on the development and 5.74% on the evaluation dataset of physical access scenario of ASVspoof 2019.

📄 PDF Abstract BibTeX arXiv:2008.03464

Code (1)

rahul-t-p/ASVspoof-2019 공식 구현

Tasks

Speaker VerificationTransfer Learning

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Residual Connection 설명 없음
Batch Normalization 설명 없음
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Average Pooling 설명 없음
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Global Average Pooling Global Average Pooling is a pooling operation designed to replace fully connected layers in classical CNNs. The idea is to generate one feature map for each corresponding…
Bottleneck Residual Block A Bottleneck Residual Block is a variant of the residual block that utilises 1x1 convolutions to create a bottleneck. The…

Similar Papers 제목 키워드 기반

Towards robust audio spoofing detection: a detailed comparison of traditional and learned features

2019-05-28 · Balamurali BT, Kin Wah Edward Lin, Simon Lui, Jer-Ming Chen 외

Automatic speaker verification, like every other biometric system, is vulnerable to spoofing attacks. Using only a few minutes of recorded voice of a genuine client of a speaker verification system, attackers can develop…

Speaker Verification

Audio Anti-spoofing Using a Simple Attention Module and Joint Optimization Based on Additive Angular Margin Loss and Meta-learning

2022-11-17 · Zhenyu Wang, John H. L. Hansen

Automatic speaker verification systems are vulnerable to a variety of access threats, prompting research into the formulation of effective spoofing detection systems to act as a gate to filter out such spoofing attacks. …

Binary ClassificationMeta-LearningSpeaker VerificationSpeech Synthesis+1

Defense for Black-box Attacks on Anti-spoofing Models by Self-Supervised Learning

2020-06-05 · Haibin Wu, Andy T. Liu, Hung-Yi Lee

High-performance anti-spoofing models for automatic speaker verification (ASV), have been widely used to protect ASV by identifying and filtering spoofing audio that is deliberately generated by text-to-speech, voice con…

Self-Supervised LearningSpeaker Verificationtext-to-speechText to Speech+1

Using Multi-Resolution Feature Maps with Convolutional Neural Networks for Anti-Spoofing in ASV

2020-08-20

This paper presents a simple but effective method that uses multi-resolution feature maps with convolutional neural networks (CNNs) for anti-spoofing in automatic speaker verification (ASV). The central idea is to allevi…

Speaker Verification

Large Audio Language Models for Spoofing-Aware Speaker Verification

2026-07-16 · Sofya Savelyeva, Mariia Perunova, Evgeny Kushnir, Artem Dvirniak 외 arxiv

Recent advances in text-to-speech and voice cloning make high-quality spoofing inexpensive and scalable, threatening voice authentication systems, especially automatic speaker verification (ASV). Existing defenses mainly…

Speaker VerificationDeepFake Detection