paper-with-me

Papers

Multi-task Metric Learning for Text-independent Speaker Verification

2020-10-21 · Yafeng Chen, Wu Guo, Jingjing Shi, Jiajun Qi, Tan Liu

In this work, we introduce metric learning (ML) to enhance the deep embedding learning for text-independent speaker verification (SV). Specifically, the deep speaker embedding network is trained with conventional cross entropy loss and auxiliary pair-based ML loss function. For the auxiliary ML task, training samples of a mini-batch are first arranged into pairs, then positive and negative pairs are selected and weighted through their own and relative similarities, and finally the auxiliary ML loss is calculated by the similarity of the selected pairs. To evaluate the proposed method, we conduct experiments on the Speaker in the Wild (SITW) dataset. The results demonstrate the effectiveness of the proposed method.

📄 PDF Abstract BibTeX arXiv:2010.10919

Code (0)

등록된 구현이 없습니다.

Tasks

Metric LearningSpeaker VerificationText-Independent Speaker Verification

Similar Papers 제목 키워드 기반

Few Shot Text-Independent speaker verification using 3D-CNN

2020-08-25 · Prateek Mishra

Facial recognition system is one of the major successes of Artificial intelligence and has been used a lot over the last years. But, images are not the only biometric present: audio is another possible biometric that can…

Speaker VerificationText-Independent Speaker Verification

Deep multi-metric learning for text-independent speaker verification

2020-07-17 · Jiwei Xu, Xinggang Wang, Bin Feng, Wenyu Liu

Text-independent speaker verification is an important artificial intelligence problem that has a wide spectrum of applications, such as criminal investigation, payment certification, and interest-based customer services.…

Metric LearningSpeaker VerificationText-Independent Speaker VerificationTriplet

Speaker-adaptive neural vocoders for parametric speech synthesis systems

2018-11-08 · Eunwoo Song, Jin-Seob Kim, Kyungguen Byun, Hong-Goo Kang

This paper proposes speaker-adaptive neural vocoders for parametric text-to-speech (TTS) systems. Recently proposed WaveNet-based neural vocoding systems successfully generate a time sequence of speech signal with an aut…

Speech Synthesistext-to-speechText to Speech

On deep speaker embeddings for text-independent speaker recognition

2018-04-26 · Sergey Novoselov, Andrey Shulipa, Ivan Kremnev, Alexandr Kozlov 외

We investigate deep neural network performance in the textindependent speaker recognition task. We demonstrate that using angular softmax activation at the last classification layer of a classification neural network ins…

General ClassificationMetric LearningSpeaker RecognitionSpeaker Verification+1

Deep Speaker Vectors for Semi Text-independent Speaker Verification

2015-05-24 · Lantian Li, Dong Wang, Zhiyong Zhang, Thomas Fang Zheng

Recent research shows that deep neural networks (DNNs) can be used to extract deep speaker vectors (d-vectors) that preserve speaker characteristics and can be used in speaker verification. This new method has been teste…

Speaker RecognitionSpeaker VerificationText-Dependent Speaker VerificationText-Independent Speaker Recognition+1