paper-with-me

홈 › Papers

Structural sparsification for Far-field Speaker Recognition with GNA

2019-10-25 · Jingchi Zhang, Jonathan Huang, Michael Deisher, Hai Li, Yiran Chen

Recently, deep neural networks (DNN) have been widely used in speaker recognition area. In order to achieve fast response time and high accuracy, the requirements for hardware resources increase rapidly. However, as the speaker recognition application is often implemented on mobile devices, it is necessary to maintain a low computational cost while keeping high accuracy in far-field condition. In this paper, we apply structural sparsification on time-delay neural networks (TDNN) to remove redundant structures and accelerate the execution. On our targeted hardware, our model can remove 60% of parameters and only slightly increasing equal error rate (EER) by 0.18% while our structural sparse model can achieve more than 1.5x speedup.

📄 PDF Abstract BibTeX arXiv:1910.11488

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Recognition

Similar Papers 제목 키워드 기반

Deep learning methods in speaker recognition: a review

2019-11-14 · Dávid Sztahó, György Szaszák, András Beke

This paper summarizes the applied deep learning practices in the field of speaker recognition, both verification and identification. Speaker recognition has been a widely used field topic of speech technology. Many resea…

Deep LearningSpeaker RecognitionSpeaker Verification

Far-Field Speaker Recognition Benchmark Derived From The DiPCo Corpus

2022-06-01 · LREC 2022 6 · Mickael Rouvier, Mohammad Mohammadamini

In this paper, we present a far-field speaker verification benchmark derived from the publicly-available DiPCo corpus. This corpus comprise three different tasks that involve enrollment and test conditions with single- a…

DenoisingSpeaker RecognitionSpeaker VerificationSpeech Enhancement+1

Multi-channel multi-speaker transformer for speech recognition

2026-01-06 · Guo Yifan, Tian Yao, Suo Hongbin, Wan Yulong arxiv

With the development of teleconferencing and in-vehicle voice assistants, far-field multi-speaker speech recognition has become a hot research topic. Recently, a multi-channel transformer (MCT) has been proposed, which d…

Speech RecognitionDeep Clustering

Phonetic-aware speaker embedding for far-field speaker verification

2023-11-27 · Zezhong Jin, Youzhi Tu, Man-Wai Mak

When a speaker verification (SV) system operates far from the sound sourced, significant challenges arise due to the interference of noise and reverberation. Studies have shown that incorporating phonetic information int…

Speaker RecognitionSpeaker Verificationspeech-recognitionSpeech Recognition

The VoxCeleb Speaker Recognition Challenge: A Retrospective

2024-08-27 · Jaesung Huh, Joon Son Chung, Arsha Nagrani, Andrew Brown 외

The VoxCeleb Speaker Recognition Challenges (VoxSRC) were a series of challenges and workshops that ran annually from 2019 to 2023. The challenges primarily evaluated the tasks of speaker recognition and diarisation unde…

Domain AdaptationSpeaker RecognitionSpeaker Verification