Scoring Formulation for Multi-Condition Joint PLDA
The joint PLDA model, is a generalization of PLDA where the nuisance variable is no longer considered independent across samples, but potentially shared (tied) across samples that correspond to the same nuisance condition. The original work considered a single nuisance condition, deriving the EM and scoring formulas for this scenario. In this document, we show how to obtain likelihood ratios for scoring when multiple nuisance conditions are allowed in the model.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Joint PLDA for Simultaneous Modeling of Two Factors
Probabilistic linear discriminant analysis (PLDA) is a method used for biometric problems like speaker or face recognition that models the variability of the samples using two latent variables, one that depends on the cl…
Face RecognitionSpeaker VerificationVocal Bursts Valence PredictionSqueezing value of cross-domain labels: a decoupled scoring approach for speaker verification
Domain mismatch often occurs in real applications and causes serious performance reduction on speaker verification systems. The common wisdom is to collect cross-domain data and train a multi-domain PLDA model, with the …
Speaker VerificationToroidal Probabilistic Spherical Discriminant Analysis
In speaker recognition, where speech segments are mapped to embeddings on the unit hypersphere, two scoring back-ends are commonly used, namely cosine scoring and PLDA. We have recently proposed PSDA, an analog to PLDA t…
FormSpeaker RecognitionJoint Speaker Encoder and Neural Back-end Model for Fully End-to-End Automatic Speaker Verification with Multiple Enrollment Utterances
Conventional automatic speaker verification systems can usually be decomposed into a front-end model such as time delay neural network (TDNN) for extracting speaker embeddings and a back-end model such as statistics-base…
Data AugmentationSpeaker VerificationProbabilistic Spherical Discriminant Analysis: An Alternative to PLDA for length-normalized embeddings
In speaker recognition, where speech segments are mapped to embeddings on the unit hypersphere, two scoring backends are commonly used, namely cosine scoring or PLDA. Both have advantages and disadvantages, depending on …
Speaker Recognition