paper-with-me

Papers

Fast variational Bayes for heavy-tailed PLDA applied to i-vectors and x-vectors

2018-03-24 · Anna Silnova, Niko Brummer, Daniel Garcia-Romero, David Snyder, Lukas Burget

The standard state-of-the-art backend for text-independent speaker recognizers that use i-vectors or x-vectors, is Gaussian PLDA (G-PLDA), assisted by a Gaussianization step involving length normalization. G-PLDA can be trained with both generative or discriminative methods. It has long been known that heavy-tailed PLDA (HT-PLDA), applied without length normalization, gives similar accuracy, but at considerable extra computational cost. We have recently introduced a fast scoring algorithm for a discriminatively trained HT-PLDA backend. This paper extends that work by introducing a fast, variational Bayes, generative training algorithm. We compare old and new backends, with and without length-normalization, with i-vectors and x-vectors, on SRE'10, SRE'16 and SITW.

📄 PDF Abstract BibTeX arXiv:1803.09153

Code (1)

bsxfan/meta-embeddings 공식 구현

Similar Papers 제목 키워드 기반

Unsupervised Adaptation of SPLDA

2015-11-20 · Jesús Villalba

State-of-the-art speaker recognition relays on models that need a large amount of training data. This models are successful in tasks like NIST SRE because there is sufficient data available. However, in real applications…

speaker-diarizationSpeaker DiarizationSpeaker Recognition

Bayesian SPLDA

2015-11-20 · Jesús Villalba

In this document we are going to derive the equations needed to implement a Variational Bayes estimation of the parameters of the simplified probabilistic linear discriminant analysis (SPLDA) model. This can be used to a…

Gaussian meta-embeddings for efficient scoring of a heavy-tailed PLDA model

2018-02-27 · Niko Brummer, Anna Silnova, Lukas Burget, Themos Stafylakis

Embeddings in machine learning are low-dimensional representations of complex input patterns, with the property that simple geometric operations like Euclidean distances and dot products can be used for classification an…

Speaker Recognition

Posterior and variational inference for deep neural networks with heavy-tailed weights

2024-06-05 · Ismaël Castillo, Paul Egels

We consider deep neural networks in a Bayesian framework with a prior distribution sampling the network weights at random. Following a recent idea of Agapiou and Castillo (2023), who show that heavy-tailed prior distribu…

Model SelectionVariational Inference

Generalized domain adaptation framework for parametric back-end in speaker recognition

2023-05-24 · Qiongqiong Wang, Koji Okabe, Kong Aik Lee, Takafumi Koshinaka

State-of-the-art speaker recognition systems comprise a speaker embedding front-end followed by a probabilistic linear discriminant analysis (PLDA) back-end. The effectiveness of these components relies on the availabili…

Domain AdaptationSpeaker RecognitionUnsupervised Domain Adaptation