UTD-CRSS Systems for 2016 NIST Speaker Recognition Evaluation
This document briefly describes the systems submitted by the Center for Robust Speech Systems (CRSS) from The University of Texas at Dallas (UTD) to the 2016 National Institute of Standards and Technology (NIST) Speaker Recognition Evaluation (SRE). We developed several UBM and DNN i-Vector based speaker recognition systems with different data sets and feature representations. Given that the emphasis of the NIST SRE 2016 is on language mismatch between training and enrollment/test data, so-called domain mismatch, in our system development we focused on: (1) using unlabeled in-domain data for centralizing data to alleviate the domain mismatch problem, (2) finding the best data set for training LDA/PLDA, (3) using newly proposed dimension reduction technique incorporating unlabeled in-domain data before PLDA training, (4) unsupervised speaker clustering of unlabeled data and using them alone or with previous SREs for PLDA training, (5) score calibration using only unlabeled data and combination of unlabeled and development (Dev) data as separate experiments.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringDimensionality ReductionSpeaker RecognitionSimilar Papers 제목 키워드 기반
Robust Speaker Recognition with Transformers Using wav2vec 2.0
Recent advances in unsupervised speech representation learning discover new approaches and provide new state-of-the-art for diverse types of speech processing tasks. This paper presents an investigation of using wav2vec …
Data AugmentationRepresentation LearningSpeaker RecognitionSpeaker Verification+1The 2021 NIST Speaker Recognition Evaluation
The 2021 Speaker Recognition Evaluation (SRE21) was the latest cycle of the ongoing evaluation series conducted by the U.S. National Institute of Standards and Technology (NIST) since 1996. It was the second large-scale …
Data AugmentationFace RecognitionPerson RecognitionSpeaker RecognitionNIST SRE CTS Superset: A large-scale dataset for telephony speaker recognition
This document provides a brief description of the National Institute of Standards and Technology (NIST) speaker recognition evaluation (SRE) conversational telephone speech (CTS) Superset. The CTS Superset has been creat…
Speaker RecognitionToeplitz Inverse Covariance based Robust Speaker Clustering for Naturalistic Audio Streams
Speaker diarization determines who spoke and when? in an audio stream. In this study, we propose a model-based approach for robust speaker clustering using i-vectors. The ivectors extracted from different segments of sam…
Clusteringspeaker-diarizationSpeaker DiarizationTHUEE system description for NIST 2019 SRE CTS Challenge
This paper describes the systems submitted by the department of electronic engineering, institute of microelectronics of Tsinghua university and TsingMicro Co. Ltd. (THUEE) to the NIST 2019 speaker recognition evaluation…
Speaker Recognition