paper-with-me

Papers

Parameterized Channel Normalization for Far-field Deep Speaker Verification

2021-09-24 · Xuechen Liu, Md Sahidullah, Tomi Kinnunen

We address far-field speaker verification with deep neural network (DNN) based speaker embedding extractor, where mismatch between enrollment and test data often comes from convolutive effects (e.g. room reverberation) and noise. To mitigate these effects, we focus on two parametric normalization methods: per-channel energy normalization (PCEN) and parameterized cepstral mean normalization (PCMN). Both methods contain differentiable parameters and thus can be conveniently integrated to, and jointly optimized with the DNN using automatic differentiation methods. We consider both fixed and trainable (data-driven) variants of each method. We evaluate the performance on Hi-MIA, a recent large-scale far-field speech corpus, with varied microphone and positional settings. Our methods outperform conventional mel filterbank features, with maximum of 33.5% and 39.5% relative improvement on equal error rate under matched microphone and mismatched microphone conditions, respectively.

📄 PDF Abstract BibTeX arXiv:2109.12056

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Self-Attentive Multi-Layer Aggregation with Feature Recalibration and Normalization for End-to-End Speaker Verification System

2020-07-27 · Soonshin Seo, Ji-Hwan Kim

One of the most important parts of an end-to-end speaker verification system is the speaker embedding generation. In our previous paper, we reported that shortcut connections-based multi-layer aggregation improves the re…

Speaker Verification

Optimized Power Normalized Cepstral Coefficients towards Robust Deep Speaker Verification

2021-09-24 · Xuechen Liu, Md Sahidullah, Tomi Kinnunen

After their introduction to robust speech recognition, power normalized cepstral coefficient (PNCC) features were successfully adopted to other tasks, including speaker verification. However, as a feature extractor with …

Robust Speech RecognitionSpeaker Verificationspeech-recognitionSpeech Recognition

MultiSV: Dataset for Far-Field Multi-Channel Speaker Verification

2021-11-11 · Ladislav Mošner, Oldřich Plchot, Lukáš Burget, Jan Černocký

Motivated by unconsolidated data situation and the lack of a standard benchmark in the field, we complement our previous efforts and present a comprehensive corpus designed for training and evaluating text-independent mu…

DenoisingSpeaker VerificationSpeech Enhancement

The 2022 Far-field Speaker Verification Challenge: Exploring domain mismatch and semi-supervised learning under the far-field scenario

2022-09-12 · Xiaoyi Qin, Ming Li, Hui Bu, Shrikanth Narayanan 외

FFSVC2022 is the second challenge of far-field speaker verification. FFSVC2022 provides the fully-supervised far-field speaker verification to further explore the far-field scenario and proposes semi-supervised far-field…

Speaker Verification

Far-Field Speaker Recognition Benchmark Derived From The DiPCo Corpus

2022-06-01 · LREC 2022 6 · Mickael Rouvier, Mohammad Mohammadamini

In this paper, we present a far-field speaker verification benchmark derived from the publicly-available DiPCo corpus. This corpus comprise three different tasks that involve enrollment and test conditions with single- a…

DenoisingSpeaker RecognitionSpeaker VerificationSpeech Enhancement+1