paper-with-me

홈 › Papers

UTD-CRSS Systems for 2016 NIST Speaker Recognition Evaluation

2016-10-24 · Chunlei Zhang, Fahimeh Bahmaninezhad, Shivesh Ranjan, Chengzhu Yu, Navid Shokouhi, John H. L. Hansen

This document briefly describes the systems submitted by the Center for Robust Speech Systems (CRSS) from The University of Texas at Dallas (UTD) to the 2016 National Institute of Standards and Technology (NIST) Speaker Recognition Evaluation (SRE). We developed several UBM and DNN i-Vector based speaker recognition systems with different data sets and feature representations. Given that the emphasis of the NIST SRE 2016 is on language mismatch between training and enrollment/test data, so-called domain mismatch, in our system development we focused on: (1) using unlabeled in-domain data for centralizing data to alleviate the domain mismatch problem, (2) finding the best data set for training LDA/PLDA, (3) using newly proposed dimension reduction technique incorporating unlabeled in-domain data before PLDA training, (4) unsupervised speaker clustering of unlabeled data and using them alone or with previous SREs for PLDA training, (5) score calibration using only unlabeled data and combination of unlabeled and development (Dev) data as separate experiments.

📄 PDF Abstract BibTeX arXiv:1610.07651

Code (0)

등록된 구현이 없습니다.

Tasks

ClusteringDimensionality ReductionSpeaker Recognition

Similar Papers 제목 키워드 기반

Robust Speaker Recognition with Transformers Using wav2vec 2.0

2022-03-28 · Sergey Novoselov, Galina Lavrentyeva, Anastasia Avdeeva, Vladimir Volokhov 외

Recent advances in unsupervised speech representation learning discover new approaches and provide new state-of-the-art for diverse types of speech processing tasks. This paper presents an investigation of using wav2vec …

Data AugmentationRepresentation LearningSpeaker RecognitionSpeaker Verification+1

The 2021 NIST Speaker Recognition Evaluation

2022-04-21 · Seyed Omid Sadjadi, Craig Greenberg, Elliot Singer, Lisa Mason 외

The 2021 Speaker Recognition Evaluation (SRE21) was the latest cycle of the ongoing evaluation series conducted by the U.S. National Institute of Standards and Technology (NIST) since 1996. It was the second large-scale …

Data AugmentationFace RecognitionPerson RecognitionSpeaker Recognition

NIST SRE CTS Superset: A large-scale dataset for telephony speaker recognition

2021-08-16 · Seyed Omid Sadjadi

This document provides a brief description of the National Institute of Standards and Technology (NIST) speaker recognition evaluation (SRE) conversational telephone speech (CTS) Superset. The CTS Superset has been creat…

Speaker Recognition

Toeplitz Inverse Covariance based Robust Speaker Clustering for Naturalistic Audio Streams

2019-07-12 · Harishchandra Dubey, Abhijeet Sangwan, John Hansen

Speaker diarization determines who spoke and when? in an audio stream. In this study, we propose a model-based approach for robust speaker clustering using i-vectors. The ivectors extracted from different segments of sam…

Clusteringspeaker-diarizationSpeaker Diarization

THUEE system description for NIST 2019 SRE CTS Challenge

2019-12-25 · Yi Liu, Tianyu Liang, Can Xu, Xianwei Zhang 외

This paper describes the systems submitted by the department of electronic engineering, institute of microelectronics of Tsinghua university and TsingMicro Co. Ltd. (THUEE) to the NIST 2019 speaker recognition evaluation…

Speaker Recognition