paper-with-me

홈 › Papers

The DKU-MSXF Speaker Verification System for the VoxCeleb Speaker Recognition Challenge 2023

2023-08-17 · Ze Li, Yuke Lin, Xiaoyi Qin, Ning Jiang, Guoqing Zhao, Ming Li

This paper is the system description of the DKU-MSXF System for the track1, track2 and track3 of the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23). For Track 1, we utilize a network structure based on ResNet for training. By constructing a cross-age QMF training set, we achieve a substantial improvement in system performance. For Track 2, we inherite the pre-trained model from Track 1 and conducte mixed training by incorporating the VoxBlink-clean dataset. In comparison to Track 1, the models incorporating VoxBlink-clean data exhibit a performance improvement by more than 10% relatively. For Track3, the semi-supervised domain adaptation task, a novel pseudo-labeling method based on triple thresholds and sub-center purification is adopted to make domain adaptation. The final submission achieves mDCF of 0.1243 in task1, mDCF of 0.1165 in Track 2 and EER of 4.952% in Track 3.

📄 PDF Abstract BibTeX arXiv:2308.08766

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationSemi-supervised Domain AdaptationSpeaker RecognitionSpeaker Verification

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Kaiming Initialization 설명 없음
Residual Connection 설명 없음
Batch Normalization 설명 없음
Bottleneck Residual Block A Bottleneck Residual Block is a variant of the residual block that utilises 1x1 convolutions to create a bottleneck. The…
Average Pooling 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

The DKU-MSXF Diarization System for the VoxCeleb Speaker Recognition Challenge 2023

2023-08-15 · Ming Cheng, Weiqing Wang, Xiaoyi Qin, Yuke Lin 외

This paper describes the DKU-MSXF submission to track 4 of the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23). Our system pipeline contains voice activity detection, clustering-based diarization, overlapped spee…

Action DetectionActivity DetectionClusteringSpeaker Recognition

Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision

2024-06-17 · Yafeng Chen, Siqi Zheng, Hui Wang, Luyao Cheng 외

Training speaker-discriminative and robust speaker verification systems without explicit speaker labels remains a persisting challenge. In this paper, we propose a new self-supervised speaker verification approach, Self-…

DiversityRepresentation LearningSpeaker VerificationTransfer Learning

Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision

2023-08-05 · Yafeng Chen, Siqi Zheng, Hui Wang, Luyao Cheng 외

Training speaker-discriminative and robust speaker verification systems without explicit speaker labels remains a persistent challenge. In this paper, we propose a novel self-supervised speaker verification approach, Sel…

DiversityRepresentation LearningSpeaker VerificationTransfer Learning

Cross-Age Speaker Verification: Learning Age-Invariant Speaker Embeddings

2022-07-13 · Xiaoyi Qin, Na Li, Chao Weng, Dan Su 외

Automatic speaker verification has achieved remarkable progress in recent years. However, there is little research on cross-age speaker verification (CASV) due to insufficient relevant data. In this paper, we mine cross-…

Age EstimationSpeaker Verification

Margin-Mixup: A Method for Robust Speaker Verification in Multi-Speaker Audio

2023-04-07 · Jenthe Thienpondt, Nilesh Madhu, Kris Demuynck

This paper is concerned with the task of speaker verification on audio with multiple overlapping speakers. Most speaker verification systems are designed with the assumption of a single speaker being present in a given a…

Speaker Verification