paper-with-me

Papers

Speaker Identification From Youtube Obtained Data

2014-11-11 · Nitesh Kumar Chaudhary

An efficient, and intuitive algorithm is presented for the identification of speakers from a long dataset (like YouTube long discussion, Cocktail party recorded audio or video).The goal of automatic speaker identification is to identify the number of different speakers and prepare a model for that speaker by extraction, characterization and speaker-specific information contained in the speech signal. It has many diverse application specially in the field of Surveillance, Immigrations at Airport, cyber security, transcription in multi-source of similar sound source, where it is difficult to assign transcription arbitrary. The most commonly speech parametrization used in speaker verification, K-mean, cepstral analysis, is detailed. Gaussian mixture modeling, which is the speaker modeling technique is then explained. Gaussian mixture models (GMM), perhaps the most robust machine learning algorithm has been introduced examine and judge carefully speaker identification in text independent. The application or employment of Gaussian mixture models for monitoring & Analysing speaker identity is encouraged by the familiarity, awareness, or understanding gained through experience that Gaussian spectrum depict the characteristics of speaker's spectral conformational pattern and remarkable ability of GMM to construct capricious densities after that we illustrate 'Expectation maximization' an iterative algorithm which takes some arbitrary value in initial estimation and carry on the iterative process until the convergence of value is observed,so by doing various number of experiments we are able to obtain 79 ~ 82% of identification rate using Vector quantization and 85 ~ 92.6% of identification rate using GMM modeling by Expectation maximization parameter estimation depending on variation of parameter.

📄 PDF Abstract BibTeX arXiv:1411.2795

Code (0)

등록된 구현이 없습니다.

Tasks

parameter estimationQuantizationSpeaker IdentificationSpeaker Verification

Similar Papers 제목 키워드 기반

Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices

2023-12-20 · Beltrán Labrador, Manuel Otero-Gonzalez, Alicia Lozano-Diez, Daniel Ramos 외

This paper presents VoxCeleb-ESP, a collection of pointers and timestamps to YouTube videos facilitating the creation of a novel speaker recognition dataset. VoxCeleb-ESP captures real-world scenarios, incorporating dive…

Speaker IdentificationSpeaker Recognition

North America Bixby Speaker Diarization System for the VoxCeleb Speaker Recognition Challenge 2021

2021-09-28 · Myungjong Kim, Taeyeon Ki, Aviral Anshu, Vijendra Raj Apsingekar

This paper describes the submission to the speaker diarization track of VoxCeleb Speaker Recognition Challenge 2021 done by North America Bixby Lab of Samsung Research America. Our speaker diarization system consists of …

Clusteringspeaker-diarizationSpeaker DiarizationSpeaker Recognition+1

Speaker Identification in each of the Neutral and Shouted Talking Environments based on Gender-Dependent Approach Using SPHMMs

2017-06-29 · Ismail Shahin

It is well known that speaker identification performs extremely well in the neutral talking environments; however, the identification performance is declined sharply in the shouted talking environments. This work aims at…

Speaker Identification

Detection and Analysis of Content Creator Collaborations in YouTube Videos using Face- and Speaker-Recognition

2018-07-05 · Moritz Lode, Michael Örtl, Christian Koch, Amr Rizk 외

This work discusses and implements the application of speaker recognition for the detection of collaborations in YouTube videos. CATANA, an existing framework for detection and analysis of YouTube collaborations, is util…

Active Speaker DetectionFace RecognitionSpeaker Recognition

VoxSRC 2022: The Fourth VoxCeleb Speaker Recognition Challenge

2023-02-20 · Jaesung Huh, Andrew Brown, Jee-weon Jung, Joon Son Chung 외

This paper summarises the findings from the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22), which was held in conjunction with INTERSPEECH 2022. The goal of this challenge was to evaluate how well state-of-the-a…

Speaker DiarizationSpeaker RecognitionSpeaker Verification