paper-with-me

홈 › Papers

Autoapprentissage pour le regroupement en locuteurs : premi\`eres investigations (First investigations on self trained speaker diarization )

2016-07-01 · JEPTALNRECITAL 2016 7 · Ga{\"e}l Le Lan, Sylvain Meignier, Delphine Charlet, Anthony Larcher

This paper investigates self trained cross-show speaker diarization applied to collections of French TV archives, based on an \textit{i-vector/PLDA} framework. The parameters used for i-vectors extraction and PLDA scoring are trained in a unsupervised way, using the data of the collection itself. Performances are compared, using combinations of target data and external data for training. The experimental results on two distinct target corpora show that using data from the corpora themselves to perform unsupervised iterative training and domain adaptation of PLDA parameters can improve an existing system, trained on external annotated data. Such results indicate that performing speaker indexation on small collections of unlabeled audio archives should only rely on the availability of a sufficient external corpus, which can be specifically adapted to every target collection. We show that a minimum collection size is required to exclude the use of such an external bootstrap.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Adaptationspeaker-diarizationSpeaker Diarization

Similar Papers 제목 키워드 기반

Nouvelle approche pour le regroupement des locuteurs dans des \'emissions radiophoniques et t\'el\'evisuelles (New approach for speaker clustering of broadcast news) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Mickael Rouvier, Sylvain Meignier
ClusteringSpeaker Diarization

Segmentation et Regroupement en Locuteurs d'une collection de documents audio (Cross-show speaker diarization) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Gr{\'e}gor Dupuy, Mickael Rouvier, Sylvain Meignier, Yannick Est{\`e}ve
speaker-diarizationSpeaker DiarizationSpeaker Verification

PTSVOX : une base de donn\'ees pour la comparaison de voix dans le cadre judiciaire (PTSVOX : a Speech Database for Forensic Voice Comparison )

2020-06-01 · JEPTALNRECITAL 2020 6 · Ana{\"\i}s Chanclu, Laurianne Georgeton, Corinne Fredouille, Jean-Francois Bonastre

Cet article pr{\'e}sente la base de donn{\'e}es PTSVOX, cr{\'e}{\'e}e par le Service Central de la Police Technique et Scientifique (SCPTS) sp{\'e}cifiquement pour la comparaison de voix dans le cadre judiciaire. PTSVOX …

Quels tests d'intelligibilit\'e pour \'evaluer les troubles de production de la parole ? (What kind of intelligibility test to assess speech production disorders?)

2016-07-01 · JEPTALNRECITAL 2016 7 · Alain Ghio, Laurence Giusti, Emilie Blanc, Serge Pinto 외

L{'}intelligibilit{\'e} de la parole se d{\'e}finit comme le degr{\'e} de pr{\'e}cision avec lequel un message est compris par un auditeur. A ce titre, la perte d{'}intelligibilit{\'e} repr{\'e}sente souvent une plainte …

R\'eseau de neurones convolutif pour l'\'evaluation automatique de la prononciation (CNN-based automatic pronunciation assessment of Japanese speakers learning French )

2016-07-01 · JEPTALNRECITAL 2016 7 · Thomas Pellegrini, Lionel Fontan, Halima Sahraoui

Dans cet article, nous comparons deux approches d{'}{\'e}valuation automatique de la prononciation de locuteurs japonophones apprenant le fran{\c{c}}ais. La premi{\`e}re, l{'}algorithme standard appel{\'e} Goodness Of Pr…