paper-with-me

Papers

Towards Robust Speaker Verification with Target Speaker Enhancement

2021-03-16 · Chunlei Zhang, Meng Yu, Chao Weng, Dong Yu

This paper proposes the target speaker enhancement based speaker verification network (TASE-SVNet), an all neural model that couples target speaker enhancement and speaker embedding extraction for robust speaker verification (SV). Specifically, an enrollment speaker conditioned speech enhancement module is employed as the front-end for extracting target speaker from its mixture with interfering speakers and environmental noises. Compared with the conventional target speaker enhancement models, nontarget speaker/interference suppression should draw additional attention for SV. Therefore, an effective nontarget speaker sampling strategy is explored. To improve speaker embedding extraction with a light-weighted model, a teacher-student (T/S) training is proposed to distill speaker discriminative information from large models to small models. Iterative inference is investigated to address the noisy speaker enrollment problem. We evaluate the proposed method on two SV tasks, i.e., one heavily overlapped speech and the other one with comprehensive noise types in vehicle environments. Experiments show significant and consistent improvements in Equal Error Rate (EER) over the state-of-the-art baselines.

📄 PDF Abstract BibTeX arXiv:2103.08781

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker VerificationSpeech Enhancement

Similar Papers 제목 키워드 기반

VoiceID Loss: Speech Enhancement for Speaker Verification

2019-04-07 · Suwon Shon, Hao Tang, James Glass

In this paper, we propose VoiceID loss, a novel loss function for training a speech enhancement model to improve the robustness of speaker verification. In contrast to the commonly used loss functions for speech enhancem…

Speaker VerificationSpeech Enhancement

Target Speaker Verification with Selective Auditory Attention for Single and Multi-talker Speech

2021-03-30 · Chenglin Xu, Wei Rao, Jibin Wu, Haizhou Li

Speaker verification has been studied mostly under the single-talker condition. It is adversely affected in the presence of interference speakers. Inspired by the study on target speaker extraction, e.g., SpEx, we propos…

Multi-Task LearningSpeaker VerificationTarget Speaker Extraction

SEF-PNet: Speaker Encoder-Free Personalized Speech Enhancement with Local and Global Contexts Aggregation

2025-01-20 · Ziling Huang, Haixin Guan, Haoran Wei, Yanhua Long

Personalized speech enhancement (PSE) methods typically rely on pre-trained speaker verification models or self-designed speaker encoders to extract target speaker clues, guiding the PSE model in isolating the desired sp…

Speaker VerificationSpeech Enhancement

Listen only to me! How well can target speech extraction handle false alarms?

2022-04-11 · Marc Delcroix, Keisuke Kinoshita, Tsubasa Ochiai, Katerina Zmolikova 외

Target speech extraction (TSE) extracts the speech of a target speaker in a mixture given auxiliary clues characterizing the speaker, such as an enrollment utterance. TSE addresses thus the challenging problem of simulta…

Speaker IdentificationSpeaker VerificationSpeech EnhancementSpeech Extraction

SASV 2022: The First Spoofing-Aware Speaker Verification Challenge

2022-03-28 · Jee-weon Jung, Hemlata Tak, Hye-jin Shim, Hee-Soo Heo 외

The first spoofing-aware speaker verification (SASV) challenge aims to integrate research efforts in speaker verification and anti-spoofing. We extend the speaker verification scenario by introducing spoofed trials to th…

Speaker Verification