paper-with-me

홈 › Papers

Siamese Neural Network with Joint Bayesian Model Structure for Speaker Verification

2021-04-07 · Xugang Lu, Peng Shen, Yu Tsao, Hisashi Kawai

Generative probability models are widely used for speaker verification (SV). However, the generative models are lack of discriminative feature selection ability. As a hypothesis test, the SV can be regarded as a binary classification task which can be designed as a Siamese neural network (SiamNN) with discriminative training. However, in most of the discriminative training for SiamNN, only the distribution of pair-wised sample distances is considered, and the additional discriminative information in joint distribution of samples is ignored. In this paper, we propose a novel SiamNN with consideration of the joint distribution of samples. The joint distribution of samples is first formulated based on a joint Bayesian (JB) based generative model, then a SiamNN is designed with dense layers to approximate the factorized affine transforms as used in the JB model. By initializing the SiamNN with the learned model parameters of the JB model, we further train the model parameters with the pair-wised samples as a binary discrimination task for SV. We carried out SV experiments on data corpus of speakers in the wild (SITW) and VoxCeleb. Experimental results showed that our proposed model improved the performance with a large margin compared with state of the art models for SV.

📄 PDF Abstract BibTeX arXiv:2104.03004

Code (0)

등록된 구현이 없습니다.

Tasks

Binary Classificationfeature selectionSpeaker Verification

Methods 이 논문이 사용한 방법론

Feature Selection Feature selection, also known as variable selection, attribute selection or variable subset selection, is the process of selecting a subset of relevant features (variables,…

Similar Papers 제목 키워드 기반

Prosodic-Enhanced Siamese Convolutional Neural Networks for Cross-Device Text-Independent Speaker Verification

2018-07-31 · Sobhan Soleymani, Ali Dabouei, Seyed Mehdi Iranmanesh, Hadi Kazemi 외

In this paper a novel cross-device text-independent speaker verification architecture is proposed. Majority of the state-of-the-art deep architectures that are used for speaker verification tasks consider Mel-frequency c…

Speaker VerificationText-Independent Speaker Verification

Coupling a generative model with a discriminative learning framework for speaker verification

2021-01-09 · Xugang Lu, Peng Shen, Yu Tsao, Hisashi Kawai

The speaker verification (SV) task is to decide whether an utterance is spoken by a target or an imposter speaker. For most studies, a log-likelihood ratio (LLR) score is estimated based on a generative probability model…

Decision Makingfeature selectionSpeaker Verification

The UPC Speaker Verification System Submitted to VoxCeleb Speaker Recognition Challenge 2020 (VoxSRC-20)

2020-10-27

This report describes the submission from Technical University of Catalonia (UPC) to the VoxCeleb Speaker Recognition Challenge (VoxSRC-20) at Interspeech 2020. The final submission is a combination of three systems. Sys…

Binary ClassificationSpeaker RecognitionSpeaker VerificationTriplet

Siamese Capsule Network for End-to-End Speaker Recognition In The Wild

2020-09-28 · Amirhossein Hajavi, Ali Etemad

We propose an end-to-end deep model for speaker verification in the wild. Our model uses thin-ResNet for extracting speaker embeddings from utterances and a Siamese capsule network and dynamic routing as the Back-end to …

Speaker RecognitionSpeaker Verification

Joint Bayesian Gaussian discriminant analysis for speaker verification

2016-12-13 · Yiyan Wang, Haotian Xu, Zhijian Ou

State-of-the-art i-vector based speaker verification relies on variants of Probabilistic Linear Discriminant Analysis (PLDA) for discriminant analysis. We are mainly motivated by the recent work of the joint Bayesian (JB…

Face VerificationSpeaker Verification