On Bottleneck Features for Text-Dependent Speaker Verification Using X-vectors
Applying x-vectors for speaker verification has recently attracted great interest, with the focus being on text-independent speaker verification. In this paper, we study x-vectors for text-dependent speaker verification (TD-SV), which remains unexplored. We further investigate the impact of the different bottleneck (BN) features on the performance of x-vectors, including the recently-introduced time-contrastive-learning (TCL) BN features and phone-discriminant BN features. TCL is a weakly supervised learning approach that constructs training data by uniformly partitioning each utterance into a predefined number of segments and then assigning each segment a class label depending on their position in the utterance. We also compare TD-SV performance for different modeling techniques, including the Gaussian mixture models-universal background model (GMM-UBM), i-vector, and x-vector. Experiments are conducted on the RedDots 2016 challenge database. It is found that the type of features has a marginal impact on the performance of x-vectors with the TCL BN feature achieving the lowest equal error rate, while the impact of features is significant for i-vector and GMM-UBM. The fusion of x-vector and i-vector systems gives a large gain in performance. The GMM-UBM technique shows its advantage for TD-SV using short utterances.
Code (0)
등록된 구현이 없습니다.
Tasks
Contrastive LearningSpeaker VerificationText-Dependent Speaker VerificationText-Independent Speaker VerificationWeakly-supervised LearningSimilar Papers 제목 키워드 기반
Time-Contrastive Learning Based DNN Bottleneck Features for Text-Dependent Speaker Verification
In this paper, we present a time-contrastive learning (TCL) based bottleneck (BN)feature extraction method for speech signals with an application to text-dependent (TD) speaker verification (SV). It is well-known that sp…
Contrastive LearningSpeaker VerificationText-Dependent Speaker VerificationComparison of Multiple Features and Modeling Methods for Text-dependent Speaker Verification
Text-dependent speaker verification is becoming popular in the speaker recognition society. However, the conventional i-vector framework which has been successful for speaker identification and other similar tasks works …
Speaker IdentificationSpeaker RecognitionSpeaker VerificationText-Dependent Speaker Verification+1End-to-End Attention based Text-Dependent Speaker Verification
A new type of End-to-End system for text-dependent speaker verification is presented in this paper. Previously, using the phonetically discriminative/speaker discriminative DNNs as feature extractors for speaker verifica…
Speaker VerificationText-Dependent Speaker VerificationTime-Contrastive Learning Based Deep Bottleneck Features for Text-Dependent Speaker Verification
There are a number of studies about extraction of bottleneck (BN) features from deep neural networks (DNNs)trained to discriminate speakers, pass-phrases and triphone states for improving the performance of text-dependen…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)ClusteringContrastive Learning+4UIAI System for Short-Duration Speaker Verification Challenge 2020
In this work, we present the system description of the UIAI entry for the short-duration speaker verification (SdSV) challenge 2020. Our focus is on Task 1 dedicated to text-dependent speaker verification. We investigate…
Speaker VerificationText-Dependent Speaker Verification