paper-with-me

홈 › Papers

A Speaker Verification Backend for Improved Calibration Performance across Varying Conditions

2020-02-05 · Luciana Ferrer, Mitchell McLaren

In a recent work, we presented a discriminative backend for speaker verification that achieved good out-of-the-box calibration performance on most tested conditions containing varying levels of mismatch to the training conditions. This backend mimics the standard PLDA-based backend process used in most current speaker verification systems, including the calibration stage. All parameters of the backend are jointly trained to optimize the binary cross-entropy for the speaker verification task. Calibration robustness is achieved by making the parameters of the calibration stage a function of vectors representing the conditions of the signal, which are extracted using a model trained to predict condition labels. In this work, we propose a simplified version of this backend where the vectors used to compute the calibration parameters are estimated within the backend, without the need for a condition prediction model. We show that this simplified method provides similar performance to the previously proposed method while being simpler to implement, and having less requirements on the training data. Further, we provide an analysis of different aspects of the method including the effect of initialization, the nature of the vectors used to compute the calibration parameters, and the effect that the random seed and the number of training epochs has on performance. We also compare the proposed method with the trial-based calibration (TBC) method that, to our knowledge, was the state-of-the-art for achieving good calibration across varying conditions. We show that the proposed method outperforms TBC while also being several orders of magnitude faster to run, comparable to the standard PLDA baseline.

📄 PDF Abstract BibTeX arXiv:2002.03802

Code (2)

Lao-Tzu-Taoism/sv_score_calibration pytorch
zh794390558/sv_score_calibration pytorch

Tasks

Speaker Verification

Similar Papers 제목 키워드 기반

A Speaker Verification Backend with Robust Performance across Conditions

2021-02-02 · Luciana Ferrer, Mitchell McLaren, Niko Brummer

In this paper, we address the problem of speaker verification in conditions unseen or unknown during development. A standard method for speaker verification consists of extracting speaker embeddings with a deep neural ne…

Speaker Verification

A discriminative condition-aware backend for speaker verification

2019-11-26 · Luciana Ferrer, Mitchell McLaren

We present a scoring approach for speaker verification that mimics the standard PLDA-based backend process used in most current speaker verification systems. However, unlike the standard backends, all parameters of the m…

Speaker Verification

Improved Relation Networks for End-to-End Speaker Verification and Identification

2022-03-31 · Ashutosh Chaubey, Sparsh Sinha, Susmita Ghose

Speaker identification systems in a real-world scenario are tasked to identify a speaker amongst a set of enrolled speakers given just a few samples for each enrolled speaker. This paper demonstrates the effectiveness of…

Meta-LearningRelationSpeaker IdentificationSpeaker Verification

Study on the Fairness of Speaker Verification Systems on Underrepresented Accents in English

2022-04-27 · Mariel Estevez, Luciana Ferrer

Speaker verification (SV) systems are currently being used to make sensitive decisions like giving access to bank accounts or deciding whether the voice of a suspect coincides with that of the perpetrator of a crime. Ens…

FairnessSpeaker Verification

NPLDA: A Deep Neural PLDA Model for Speaker Verification

2020-02-10 · Shreyas Ramoji, Prashant Krishnan, Sriram Ganapathy

The state-of-art approach for speaker verification consists of a neural network based embedding extractor along with a backend generative model such as the Probabilistic Linear Discriminant Analysis (PLDA). In this work,…

Speaker RecognitionSpeaker Verification