paper-with-me

Papers

Data augmentation versus noise compensation for x- vector speaker recognition systems in noisy environments

2020-06-29 · Mohammad Mohammadamini, Driss Matrouf

The explosion of available speech data and new speaker modeling methods based on deep neural networks (DNN) have given the ability to develop more robust speaker recognition systems. Among DNN speaker modelling techniques, x-vector system has shown a degree of robustness in noisy environments. Previous studies suggest that by increasing the number of speakers in the training data and using data augmentation more robust speaker recognition systems are achievable in noisy environments. In this work, we want to know if explicit noise compensation techniques continue to be effective despite the general noise robustness of these systems. For this study, we will use two different x-vector networks: the first one is trained on Voxceleb1 (Protocol1), and the second one is trained on Voxceleb1+Voxveleb2 (Protocol2). We propose to add a denoising x-vector subsystem before scoring. Experimental results show that, the x-vector system used in Protocol2 is more robust than the other one used Protocol1. Despite this observation we will show that explicit noise compensation gives almost the same EER relative gain in both protocols. For example, in the Protocol2 we have 21% to 66% improvement of EER with denoising techniques.

📄 PDF Abstract BibTeX arXiv:2006.15903

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationDenoisingSpeaker Recognition

Similar Papers 제목 키워드 기반

Enhancing 5G-NR mmWave : Phase Noise Models Evaluation with MMSE for CPE Compensation

2024-12-08 · Desire Guel, Flavien Herve Somda, Boureima Zerbo, Oumarou Sie

The rapid development of 5G New Radio (NR) and millimeter-wave (mmWave) communication systems highlights the critical importance of maintaining accurate phase synchronization to ensure reliable and efficient communicatio…

Hearing-Loss Compensation Using Deep Neural Networks: A Framework and Results From a Listening Test

2024-03-15 · Peter Leer, Jesper Jensen, Laurel H. Carney, Zheng-Hua Tan 외

This article investigates the use of deep neural networks (DNNs) for hearing-loss compensation. Hearing loss is a prevalent issue affecting millions of people worldwide, and conventional hearing aids have limitations in …

Music ClassificationSpeaker Identificationspeech-recognitionSpeech Recognition

Vocoder drift compensation by x-vector alignment in speaker anonymisation

2023-07-17 · Michele Panariello, Massimiliano Todisco, Nicholas Evans

For the most popular x-vector-based approaches to speaker anonymisation, the bulk of the anonymisation can stem from vocoding rather than from the core anonymisation function which is used to substitute an original speak…

Speech Enhancement in Adverse Environments Based on Non-stationary Noise-driven Spectral Subtraction and SNR-dependent Phase Compensation

2018-02-19

A two-step enhancement method based on spectral subtraction and phase spectrum compensation is presented in this paper for noisy speeches in adverse environments involving non-stationary noise and medium to low levels of…

Noise EstimationSpeech Enhancement

Augmentation Strategies for Learning with Noisy Labels

2021-03-03 · CVPR 2021 1 · Kento Nishi, Yi Ding, Alex Rich, Tobias Höllerer

Imperfect labels are ubiquitous in real-world datasets. Several recent successful methods for training deep neural networks (DNNs) robust to label noise have used two primary techniques: filtering samples based on loss d…

Image ClassificationLearning with noisy labelsPseudo Label