Weakly Supervised PLDA Training
PLDA is a popular normalization approach for the i-vector model, and it has delivered state-of-the-art performance in speaker verification. However, PLDA training requires a large amount of labelled development data, which is highly expensive in most cases. We present a cheap PLDA training approach, which assumes that speakers in the same session can be easily separated, and speakers in different sessions are simply different. This results in `weak labels' which are not fully accurate but cheap, leading to a weak PLDA training. Our experimental results on real-life large-scale telephony customer service achieves demonstrated that the weak training can offer good performance when human-labelled data are limited. More interestingly, the weak training can be employed as a discriminative adaptation approach, which is more efficient than the prevailing unsupervised method when human-labelled data are insufficient.
Code (0)
등록된 구현이 없습니다.
Tasks
Speaker VerificationSimilar Papers 제목 키워드 기반
Unsupervised Adaptation of SPLDA
State-of-the-art speaker recognition relays on models that need a large amount of training data. This models are successful in tasks like NIST SRE because there is sufficient data available. However, in real applications…
speaker-diarizationSpeaker DiarizationSpeaker RecognitionMooseNet: A Trainable Metric for Synthesized Speech with a PLDA Module
We present MooseNet, a trainable speech metric that predicts the listeners' Mean Opinion Score (MOS). We propose a novel approach where the Probabilistic Linear Discriminative Analysis (PLDA) generative model is used on …
Self-Supervised LearningLocal Training for PLDA in Speaker Verification
PLDA is a popular normalization approach for the i-vector model, and it has delivered state-of-the-art performance in speaker verification. However, PLDA training requires a large amount of labeled development data, whic…
Speaker VerificationGeneralized domain adaptation framework for parametric back-end in speaker recognition
State-of-the-art speaker recognition systems comprise a speaker embedding front-end followed by a probabilistic linear discriminant analysis (PLDA) back-end. The effectiveness of these components relies on the availabili…
Domain AdaptationSpeaker RecognitionUnsupervised Domain AdaptationAutoapprentissage pour le regroupement en locuteurs : premi\`eres investigations (First investigations on self trained speaker diarization )
This paper investigates self trained cross-show speaker diarization applied to collections of French TV archives, based on an \textit{i-vector/PLDA} framework. The parameters used for i-vectors extraction and PLDA scorin…
Domain Adaptationspeaker-diarizationSpeaker Diarization