VAE-based Domain Adaptation for Speaker Verification
Deep speaker embedding has achieved satisfactory performance in speaker verification. By enforcing the neural model to discriminate the speakers in the training set, deep speaker embedding (called x-vectors) can be derived from the hidden layers. Despite its good performance, the present embedding model is highly domain sensitive, which means that it often works well in domains whose acoustic condition matches that of the training data (in-domain), but degrades in mismatched domains (out-of-domain). In this paper, we present a domain adaptation approach based on Variational Auto-Encoder (VAE). This model transforms x-vectors to a regularized latent space; within this latent space, a small amount of data from the target domain is sufficient to accomplish the adaptation. Our experiments demonstrated that by this VAE-adaptation approach, speaker embeddings can be easily transformed to the target domain, leading to noticeable performance improvement.
Code (0)
등록된 구현이 없습니다.
Tasks
Domain AdaptationSpeaker VerificationSimilar Papers 제목 키워드 기반
Cross-lingual Text-independent Speaker Verification using Unsupervised Adversarial Discriminative Domain Adaptation
Speaker verification systems often degrade significantly when there is a language mismatch between training and testing data. Being able to improve cross-lingual speaker verification system using unlabeled data can great…
Domain AdaptationSpeaker VerificationText-Independent Speaker VerificationSource -Free Domain Adaptation for Speaker Verification in Data-Scarce Languages and Noisy Channels
Domain adaptation is often hampered by exceedingly small target datasets and inaccessible source data. These conditions are prevalent in speech verification, where privacy policies and/or languages with scarce speech res…
Domain AdaptationSource-Free Domain AdaptationSpeaker VerificationDeep Feature CycleGANs: Speaker Identity Preserving Non-parallel Microphone-Telephone Domain Adaptation for Speaker Verification
With the increase in the availability of speech from varied domains, it is imperative to use such out-of-domain data to improve existing speech systems. Domain adaptation is a prominent pre-processing approach for this. …
Domain AdaptationSpeaker VerificationTranslationPrototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
Speaker verification system trained on one domain usually suffers performance degradation when applied to another domain. To address this challenge, researchers commonly use feature distribution matching-based methods in…
Contrastive LearningDomain AdaptationSpeaker VerificationUnsupervised Domain AdaptationCross-domain Adaptation with Discrepancy Minimization for Text-independent Forensic Speaker Verification
Forensic audio analysis for speaker verification offers unique challenges due to location/scenario uncertainty and diversity mismatch between reference and naturalistic field recordings. The lack of real naturalistic for…
DiversityDomain AdaptationSpeaker Verification