paper-with-me

Papers

Max-margin Metric Learning for Speaker Recognition

2015-10-20 · Lantian Li, Dong Wang, Chao Xing, Thomas Fang Zheng

Probabilistic linear discriminant analysis (PLDA) is a popular normalization approach for the i-vector model, and has delivered state-of-the-art performance in speaker recognition. A potential problem of the PLDA model, however, is that it essentially assumes Gaussian distributions over speaker vectors, which is not always true in practice. Additionally, the objective function is not directly related to the goal of the task, e.g., discriminating true speakers and imposters. In this paper, we propose a max-margin metric learning approach to solve the problems. It learns a linear transform with a criterion that the margin between target and imposter trials are maximized. Experiments conducted on the SRE08 core test show that compared to PLDA, the new approach can obtain comparable or even better performance, though the scoring is simply a cosine computation.

📄 PDF Abstract BibTeX arXiv:1510.05940

Code (0)

등록된 구현이 없습니다.

Tasks

Metric LearningSpeaker Recognition

Similar Papers 제목 키워드 기반

Curricular SincNet: Towards Robust Deep Speaker Recognition by Emphasizing Hard Samples in Latent Space

2021-08-21 · Labib Chowdhury, Mustafa Kamal, Najia Hasan, Nabeel Mohammed

Deep learning models have become an increasingly preferred option for biometric recognition systems, such as speaker recognition. SincNet, a deep neural network architecture, gained popularity in speaker recognition task…

Face RecognitionSpeaker Recognition

Challenging margin-based speaker embedding extractors by using the variational information bottleneck

2024-06-18 · Themos Stafylakis, Anna Silnova, Johan Rohdin, Oldrich Plchot 외

Speaker embedding extractors are typically trained using a classification loss over the training speakers. During the last few years, the standard softmax/cross-entropy loss has been replaced by the margin-based losses, …

Speaker Recognition

Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition

2019-06-18 · Xu Xiang, Shuai Wang, Houjun Huang, Yanmin Qian 외

Recently, speaker embeddings extracted from a speaker discriminative deep neural network (DNN) yield better performance than the conventional methods such as i-vector. In most cases, the DNN speaker classifier is trained…

Speaker Recognition

Adversarial defense for deep speaker recognition using hybrid adversarial training

2020-10-30

Deep neural network based speaker recognition systems can easily be deceived by an adversary using minuscule imperceptible perturbations to the input speech samples. These adversarial attacks pose serious security threat…

Adversarial DefenseSpeaker Recognition

VoxCeleb2: Deep Speaker Recognition

2018-06-14 · Joon Son Chung, Arsha Nagrani, Andrew Zisserman

The objective of this paper is speaker recognition under noisy and unconstrained conditions. We make two key contributions. First, we introduce a very large-scale audio-visual speaker recognition dataset collected from…

Speaker RecognitionSpeaker Verification