paper-with-me

홈 › Papers

A Speaker Verification Backend with Robust Performance across Conditions

2021-02-02 · Luciana Ferrer, Mitchell McLaren, Niko Brummer

In this paper, we address the problem of speaker verification in conditions unseen or unknown during development. A standard method for speaker verification consists of extracting speaker embeddings with a deep neural network and processing them through a backend composed of probabilistic linear discriminant analysis (PLDA) and global logistic regression score calibration. This method is known to result in systems that work poorly on conditions different from those used to train the calibration model. We propose to modify the standard backend, introducing an adaptive calibrator that uses duration and other automatically extracted side-information to adapt to the conditions of the inputs. The backend is trained discriminatively to optimize binary cross-entropy. When trained on a number of diverse datasets that are labeled only with respect to speaker, the proposed backend consistently and, in some cases, dramatically improves calibration, compared to the standard PLDA approach, on a number of held-out datasets, some of which are markedly different from the training data. Discrimination performance is also consistently improved. We show that joint training of the PLDA and the adaptive calibrator is essential -- the same benefits cannot be achieved when freezing PLDA and fine-tuning the calibrator. To our knowledge, the results in this paper are the first evidence in the literature that it is possible to develop a speaker verification system with robust out-of-the-box performance on a large variety of conditions.

📄 PDF Abstract BibTeX arXiv:2102.01760

Code (1)

luferrer/DCA-PLDA 공식 구현 pytorch

Tasks

Speaker Verification

Methods 이 논문이 사용한 방법론

Logistic Regression Logistic Regression, despite its name, is a linear model for classification rather than regression. Logistic regression is also known in the literature as logit regression,…

Similar Papers 제목 키워드 기반

A Speaker Verification Backend for Improved Calibration Performance across Varying Conditions

2020-02-05 · Luciana Ferrer, Mitchell McLaren

In a recent work, we presented a discriminative backend for speaker verification that achieved good out-of-the-box calibration performance on most tested conditions containing varying levels of mismatch to the training c…

Speaker Verification

A discriminative condition-aware backend for speaker verification

2019-11-26 · Luciana Ferrer, Mitchell McLaren

We present a scoring approach for speaker verification that mimics the standard PLDA-based backend process used in most current speaker verification systems. However, unlike the standard backends, all parameters of the m…

Speaker Verification

Can spoofing countermeasure and speaker verification systems be jointly optimised?

2023-03-13 · Wanying Ge, Hemlata Tak, Massimiliano Todisco, Nicholas Evans

Spoofing countermeasure (CM) and automatic speaker verification (ASV) sub-systems can be used in tandem with a backend classifier as a solution to the spoofing aware speaker verification (SASV) task. The two sub-systems …

Speaker Verification

NPLDA: A Deep Neural PLDA Model for Speaker Verification

2020-02-10 · Shreyas Ramoji, Prashant Krishnan, Sriram Ganapathy

The state-of-art approach for speaker verification consists of a neural network based embedding extractor along with a backend generative model such as the Probabilistic Linear Discriminant Analysis (PLDA). In this work,…

Speaker RecognitionSpeaker Verification

Generalizing Speaker Verification for Spoof Awareness in the Embedding Space

2024-01-20 · Xuechen Liu, Md Sahidullah, Kong Aik Lee, Tomi Kinnunen

It is now well-known that automatic speaker verification (ASV) systems can be spoofed using various types of adversaries. The usual approach to counteract ASV systems against such attacks is to develop a separate spoofin…

Domain AdaptationSpeaker Verification