PLDA with Two Sources of Inter-session Variability
In some speaker recognition scenarios we find conversations recorded simultaneously over multiple channels. That is the case of the interviews in the NIST SRE dataset. To take advantage of that, we propose a modification of the PLDA model that considers two different inter-session variability terms. The first term is tied between all the recordings belonging to the same conversation whereas the second is not. Thus, the former mainly intends to capture the variability due to the phonetic content of the conversation while the latter tries to capture the channel variability. In this document, we derive the equations for this model. This model was applied in the paper "Handling Recordings Acquired Simultaneously over Multiple Channels with PLDA" published at Interspeech 2013.
Code (0)
등록된 구현이 없습니다.
Tasks
Speaker RecognitionVocal Bursts Valence PredictionSimilar Papers 제목 키워드 기반
Weakly Supervised PLDA Training
PLDA is a popular normalization approach for the i-vector model, and it has delivered state-of-the-art performance in speaker verification. However, PLDA training requires a large amount of labelled development data, whi…
Speaker VerificationJoint PLDA for Simultaneous Modeling of Two Factors
Probabilistic linear discriminant analysis (PLDA) is a method used for biometric problems like speaker or face recognition that models the variability of the samples using two latent variables, one that depends on the cl…
Face RecognitionSpeaker VerificationVocal Bursts Valence PredictionJoint Probabilistic Linear Discriminant Analysis
Standard probabilistic linear discriminant analysis (PLDA) for speaker recognition assumes that the sample's features (usually, i-vectors) are given by a sum of three terms: a term that depends on the speaker identity, a…
Speaker RecognitionVariable frame rate-based data augmentation to handle speaking-style variability for automatic speaker verification
The effects of speaking-style variability on automatic speaker verification were investigated using the UCLA Speaker Variability database which comprises multiple speaking styles per speaker. An x-vector/PLDA (probabilis…
Data AugmentationSpeaker VerificationMinimizing inter-subject variability in fNIRS based Brain Computer Interfaces via multiple-kernel support vector learning
Brain signal variability in the measurements obtained from different subjects during different sessions significantly deteriorates the accuracy of most brain-computer interface (BCI) systems. Moreover these variabilities…
Brain Computer Interface