paper-with-me

홈 › Papers

A Multi Level Data Fusion Approach for Speaker Identification on Telephone Speech

2014-06-27 · Imen Trabelsi, Dorra Ben Ayed

Several speaker identification systems are giving good performance with clean speech but are affected by the degradations introduced by noisy audio conditions. To deal with this problem, we investigate the use of complementary information at different levels for computing a combined match score for the unknown speaker. In this work, we observe the effect of two supervised machine learning approaches including support vectors machines (SVM) and na\"ive bayes (NB). We define two feature vector sets based on mel frequency cepstral coefficients (MFCC) and relative spectral perceptual linear predictive coefficients (RASTA-PLP). Each feature is modeled using the Gaussian Mixture Model (GMM). Several ways of combining these information sources give significant improvements in a text-independent speaker identification task using a very large telephone degraded NTIMIT database.

📄 PDF Abstract BibTeX arXiv:1407.0380

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Identification

Similar Papers 제목 키워드 기반

Graph-based Multi-View Fusion and Local Adaptation: Mitigating Within-Household Confusability for Speaker Identification

2022-07-08 · Long Chen, Yixiong Meng, Venkatesh Ravichandran, Andreas Stolcke

Speaker identification (SID) in the household scenario (e.g., for smart speakers) is an important but challenging problem due to limited number of labeled (enrollment) utterances, confusable voices, and demographic imbal…

FairnessSpeaker IdentificationSpeaker Recognition

Weakly Supervised Training of Speaker Identification Models

2018-06-22 · Martin Karu, Tanel Alumäe

We propose an approach for training speaker identification models in a weakly supervised manner. We concentrate on the setting where the training data consists of a set of audio recordings and the speaker annotation is p…

speaker-diarizationSpeaker DiarizationSpeaker Identification

Weakly Supervised Training of Hierarchical Attention Networks for Speaker Identification

2020-05-15 · Yanpei Shi, Qiang Huang, Thomas Hain

Identifying multiple speakers without knowing where a speaker's voice is in a recording is a challenging task. In this paper, a hierarchical attention network is proposed to solve a weakly labelled speaker identification…

Speaker Identification

AMR: Adaptive Modality Routing for Multimodal Polyglot Speaker Identification

2026-06-28 · Chuxiao Zuo, Yao Zhu, Minqiang Xu, Manhong Wang 외 arxiv

Multimodal speaker identification systems face two key challenges in real-world deployment: missing modalities and language mismatch between training and testing conditions. In practical scenarios, background multi-speak…

Speaker Identification

Latent space representation for multi-target speaker detection and identification with a sparse dataset using Triplet neural networks

2019-10-01 · Kin Wai Cheuk, Balamurali B. T., Gemma Roig, Dorien Herremans

We present an approach to tackle the speaker recognition problem using Triplet Neural Networks. Currently, the $i$-vector representation with probabilistic linear discriminant analysis (PLDA) is the most commonly used te…

Speaker IdentificationSpeaker RecognitionTriplet