paper-with-me

홈 › Papers

The HUAWEI Speaker Diarisation System for the VoxCeleb Speaker Diarisation Challenge

2020-10-22 · Renyu Wang, Ruilin Tong, Yu Ting Yeung, Xiao Chen

This paper describes system setup of our submission to speaker diarisation track (Track 4) of VoxCeleb Speaker Recognition Challenge 2020. Our diarisation system consists of a well-trained neural network based speech enhancement model as pre-processing front-end of input speech signals. We replace conventional energy-based voice activity detection (VAD) with a neural network based VAD. The neural network based VAD provides more accurate annotation of speech segments containing only background music, noise, and other interference, which is crucial to diarisation performance. We apply agglomerative hierarchical clustering (AHC) of x-vectors and variational Bayesian hidden Markov model (VB-HMM) based iterative clustering for speaker clustering. Experimental results demonstrate that our proposed system achieves substantial improvements over the baseline system, yielding diarisation error rate (DER) of 10.45%, and Jacard error rate (JER) of 22.46% on the evaluation set.

📄 PDF Abstract BibTeX arXiv:2010.11657

Code (0)

등록된 구현이 없습니다.

Tasks

Action DetectionActivity DetectionClusteringSpeaker RecognitionSpeech Enhancement

Similar Papers 제목 키워드 기반

XMUSPEECH System for VoxCeleb Speaker Recognition Challenge 2021

2021-09-06 · Jie Wang, Fuchuang Tong, Zhicong Chen, Lin Li 외

This paper describes the XMUSPEECH speaker recognition and diarisation systems for the VoxCeleb Speaker Recognition Challenge 2021. For track 2, we evaluate two systems including ResNet34-SE and ECAPA-TDNN. For track 4, …

Speaker Recognition

The VoxCeleb Speaker Recognition Challenge: A Retrospective

2024-08-27 · Jaesung Huh, Joon Son Chung, Arsha Nagrani, Andrew Brown 외

The VoxCeleb Speaker Recognition Challenges (VoxSRC) were a series of challenges and workshops that ran annually from 2019 to 2023. The challenges primarily evaluated the tasks of speaker recognition and diarisation unde…

Domain AdaptationSpeaker RecognitionSpeaker Verification

VoxSRC 2022: The Fourth VoxCeleb Speaker Recognition Challenge

2023-02-20 · Jaesung Huh, Andrew Brown, Jee-weon Jung, Joon Son Chung 외

This paper summarises the findings from the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22), which was held in conjunction with INTERSPEECH 2022. The goal of this challenge was to evaluate how well state-of-the-a…

Speaker DiarizationSpeaker RecognitionSpeaker Verification

VoxSRC 2020: The Second VoxCeleb Speaker Recognition Challenge

2020-12-12 · Arsha Nagrani, Joon Son Chung, Jaesung Huh, Andrew Brown 외

We held the second installment of the VoxCeleb Speaker Recognition Challenge in conjunction with Interspeech 2020. The goal of this challenge was to assess how well current speaker recognition technology is able to diari…

Speaker Recognition

Combination of Deep Speaker Embeddings for Diarisation

2020-10-22 · Guangzhi Sun, Chao Zhang, Phil Woodland

Significant progress has recently been made in speaker diarisation after the introduction of d-vectors as speaker embeddings extracted from neural network (NN) speaker classifiers for clustering speech segments. To extra…

Action DetectionActivity DetectionChange Point DetectionClustering