paper-with-me

Papers

Powerful Speaker Embedding Training Framework by Adversarially Disentangled Identity Representation

2019-11-27 · Jianwei Tai, Hang Zhou, Qingjia Huang, Xiaoqi Jia

The main challenge of speaker verification in the wild is the interference caused by irrelevant information in speech and the lack of speaker labels in speech datasets. In order to solve the above problems, we propose a novel speaker embedding training framework based on adversarially disentangled identity representation. Our key insight is to adversarially learn the identity-purified features for speaker verification, and learn an identity-irrelated feature whose speaker information cannot be distinguished. Based on the existing state-of-the-art speaker verification models, we improve them without adjusting the structure and hyper-parameters of any model. Experiments prove that the framework we propose can significantly improve the performance of speaker verification from the original model without any empirical adjustments. Proving that it is particularly useful for alleviating the lack of speaker labels.

📄 PDF Abstract BibTeX arXiv:1912.02608

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Similar Papers 제목 키워드 기반

Adversarial Attacks and Robust Defenses in Speaker Embedding based Zero-Shot Text-to-Speech System

2024-10-05 · Ze Li, Yao Shi, Yunfei Xu, Ming Li

Speaker embedding based zero-shot Text-to-Speech (TTS) systems enable high-quality speech synthesis for unseen speakers using minimal data. However, these systems are vulnerable to adversarial attacks, where an attacker …

Adversarial PurificationSpeech Synthesistext-to-speechText to Speech

Channel adversarial training for speaker verification and diarization

2019-10-25 · Chau Luu, Peter Bell, Steve Renals

Previous work has encouraged domain-invariance in deep speaker embedding by adversarially classifying the dataset or labelled environment to which the generated features belong. We propose a training strategy which aims …

Speaker Verification

Deep Representation Decomposition for Rate-Invariant Speaker Verification

2022-05-28 · Fuchuan Tong, Siqi Zheng, Haodong Zhou, Xingjia Xie 외

While promising performance for speaker verification has been achieved by deep speaker embeddings, the advantage would reduce in the case of speaking-style variability. Speaking rate mismatch is often observed in practic…

Speaker Verification

Multi-target Voice Conversion without Parallel Data by Adversarially Learning Disentangled Audio Representations

2018-04-09 · Ju-chieh Chou, Cheng-chieh Yeh, Hung-Yi Lee, Lin-shan Lee

Recently, cycle-consistent adversarial network (Cycle-GAN) has been successfully applied to voice conversion to a different speaker without parallel data, although in those approaches an individual model is needed for ea…

DecoderVoice Conversion

ECAPA-TDNN Embeddings for Speaker Diarization

2021-04-03 · Nauman Dawalatabad, Mirco Ravanelli, François Grondin, Jenthe Thienpondt 외

Learning robust speaker embeddings is a crucial step in speaker diarization. Deep neural networks can accurately capture speaker discriminative characteristics and popular deep embeddings such as x-vectors are nowadays a…

speaker-diarizationSpeaker DiarizationSpeaker Verification