paper-with-me

Papers

Far-Field Speaker Recognition Benchmark Derived From The DiPCo Corpus

2022-06-01 · LREC 2022 6 · Mickael Rouvier, Mohammad Mohammadamini

In this paper, we present a far-field speaker verification benchmark derived from the publicly-available DiPCo corpus. This corpus comprise three different tasks that involve enrollment and test conditions with single- and/or multi-channels recordings. The main goal of this corpus is to foster research in far-field and multi-channel text-independent speaker verification. Also, it can be used for other speaker recognition tasks such as dereverberation, denoising and speech enhancement. In addition, we release a Kaldi and SpeechBrain system to facilitate further research. And we validate the evaluation design with a single-microphone state-of-the-art speaker recognition system (i.e. ResNet-101). The results show that the proposed tasks are very challenging. And we hope these resources will inspire the speech community to develop new methods and systems for this challenging domain.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingSpeaker RecognitionSpeaker VerificationSpeech EnhancementText-Independent Speaker Verification

Similar Papers 제목 키워드 기반

Exploring Speaker Diarization with Mixture of Experts

2025-06-17 · Gaobin Yang, Maokui He, Shutong Niu, Ruoyu Wang 외

In this paper, we propose a novel neural speaker diarization system using memory-aware multi-speaker embedding with sequence-to-sequence architecture (NSD-MS2S), which integrates a memory-aware multi-speaker embedding mo…

Mixture-of-Expertsspeaker-diarizationSpeaker Diarization

DiPCo -- Dinner Party Corpus

2019-09-30 · Maarten Van Segbroeck, Ahmed Zaid, Ksenia Kutsenko, Cirenia Huerta 외

We present a speech data corpus that simulates a "dinner party" scenario taking place in an everyday home environment. The corpus was created by recording multiple groups of four Amazon employee volunteers having a natur…

Benchmarking

AutoSpeech: Neural Architecture Search for Speaker Recognition

2020-05-07 · Shaojin Ding, Tianlong Chen, Xinyu Gong, Weiwei Zha 외

Speaker recognition systems based on Convolutional Neural Networks (CNNs) are often built with off-the-shelf backbones such as VGG-Net or ResNet. However, these backbones were originally proposed for image classification…

image-classificationImage ClassificationNeural Architecture SearchSpeaker Identification+2

Exploring Multilingual Unseen Speaker Emotion Recognition: Leveraging Co-Attention Cues in Multitask Learning

2024-06-13 · Arnav Goel, Medha Hira, Anubha Gupta

Advent of modern deep learning techniques has given rise to advancements in the field of Speech Emotion Recognition (SER). However, most systems prevalent in the field fail to generalize to speakers not seen during train…

Emotion RecognitionSpeech Emotion Recognition

Baselines and Protocols for Household Speaker Recognition

2022-04-30 · Alexey Sholokhov, Xuechen Liu, Md Sahidullah, Tomi Kinnunen

Speaker recognition on household devices, such as smart speakers, features several challenges: (i) robustness across a vast number of heterogeneous domains (households), (ii) short utterances, (iii) possibly absent speak…

Speaker Recognition