Multi-task Metric Learning for Text-independent Speaker Verification
In this work, we introduce metric learning (ML) to enhance the deep embedding learning for text-independent speaker verification (SV). Specifically, the deep speaker embedding network is trained with conventional cross entropy loss and auxiliary pair-based ML loss function. For the auxiliary ML task, training samples of a mini-batch are first arranged into pairs, then positive and negative pairs are selected and weighted through their own and relative similarities, and finally the auxiliary ML loss is calculated by the similarity of the selected pairs. To evaluate the proposed method, we conduct experiments on the Speaker in the Wild (SITW) dataset. The results demonstrate the effectiveness of the proposed method.
Code (0)
등록된 구현이 없습니다.
Tasks
Metric LearningSpeaker VerificationText-Independent Speaker VerificationSimilar Papers 제목 키워드 기반
Few Shot Text-Independent speaker verification using 3D-CNN
Facial recognition system is one of the major successes of Artificial intelligence and has been used a lot over the last years. But, images are not the only biometric present: audio is another possible biometric that can…
Speaker VerificationText-Independent Speaker VerificationDeep multi-metric learning for text-independent speaker verification
Text-independent speaker verification is an important artificial intelligence problem that has a wide spectrum of applications, such as criminal investigation, payment certification, and interest-based customer services.…
Metric LearningSpeaker VerificationText-Independent Speaker VerificationTripletSpeaker-adaptive neural vocoders for parametric speech synthesis systems
This paper proposes speaker-adaptive neural vocoders for parametric text-to-speech (TTS) systems. Recently proposed WaveNet-based neural vocoding systems successfully generate a time sequence of speech signal with an aut…
Speech Synthesistext-to-speechText to SpeechOn deep speaker embeddings for text-independent speaker recognition
We investigate deep neural network performance in the textindependent speaker recognition task. We demonstrate that using angular softmax activation at the last classification layer of a classification neural network ins…
General ClassificationMetric LearningSpeaker RecognitionSpeaker Verification+1Deep Speaker Vectors for Semi Text-independent Speaker Verification
Recent research shows that deep neural networks (DNNs) can be used to extract deep speaker vectors (d-vectors) that preserve speaker characteristics and can be used in speaker verification. This new method has been teste…
Speaker RecognitionSpeaker VerificationText-Dependent Speaker VerificationText-Independent Speaker Recognition+1