paper-with-me

Papers

Robust speaker recognition using unsupervised adversarial invariance

2019-11-03 · Raghuveer Peri, Monisankha Pal, Arindam Jati, Krishna Somandepalli, Shrikanth Narayanan

In this paper, we address the problem of speaker recognition in challenging acoustic conditions using a novel method to extract robust speaker-discriminative speech representations. We adopt a recently proposed unsupervised adversarial invariance architecture to train a network that maps speaker embeddings extracted using a pre-trained model onto two lower dimensional embedding spaces. The embedding spaces are learnt to disentangle speaker-discriminative information from all other information present in the audio recordings, without supervision about the acoustic conditions. We analyze the robustness of the proposed embeddings to various sources of variability present in the signal for speaker verification and unsupervised clustering tasks on a large-scale speaker recognition corpus. Our analyses show that the proposed system substantially outperforms the baseline in a variety of challenging acoustic scenarios. Furthermore, for the task of speaker diarization on a real-world meeting corpus, our system shows a relative improvement of 36\% in the diarization error rate compared to the state-of-the-art baseline.

📄 PDF Abstract BibTeX arXiv:1911.00940

Code (1)

rperi/speaker-embeddings-UAI-inference tf

Tasks

speaker-diarizationSpeaker DiarizationSpeaker RecognitionSpeaker Verification

Similar Papers 제목 키워드 기반

To Reverse the Gradient or Not: An Empirical Comparison of Adversarial and Multi-task Learning in Speech Recognition

2018-12-09 · Yossi Adi, Neil Zeghidour, Ronan Collobert, Nicolas Usunier 외

Transcribed datasets typically contain speaker identity for each instance in the data. We investigate two ways to incorporate this information during training: Multi-Task Learning and Adversarial Learning. In multi-task …

Multi-Task LearningSpeaker Recognitionspeech-recognitionSpeech Recognition

Adversarial defense for deep speaker recognition using hybrid adversarial training

2020-10-30

Deep neural network based speaker recognition systems can easily be deceived by an adversary using minuscule imperceptible perturbations to the input speech samples. These adversarial attacks pose serious security threat…

Adversarial DefenseSpeaker Recognition

NIESR: Nuisance Invariant End-to-end Speech Recognition

2019-07-07 · I-Hung Hsu, Ayush Jaiswal, Premkumar Natarajan

Deep neural network models for speech recognition have achieved great success recently, but they can learn incorrect associations between the target and nuisance factors of speech (e.g., speaker identities, background no…

speech-recognitionSpeech Recognition

Augmentation adversarial training for self-supervised speaker recognition

2020-07-23 · Jaesung Huh, Hee Soo Heo, Jingu Kang, Shinji Watanabe 외

The goal of this work is to train robust speaker recognition models without speaker labels. Recent works on unsupervised speaker representations are based on contrastive learning in which they encourage within-utterance …

Contrastive LearningSpeaker Recognition

Channel adversarial training for speaker verification and diarization

2019-10-25 · Chau Luu, Peter Bell, Steve Renals

Previous work has encouraged domain-invariance in deep speaker embedding by adversarially classifying the dataset or labelled environment to which the generated features belong. We propose a training strategy which aims …

Speaker Verification