The ETAPE speech processing evaluation
The ETAPE evaluation is the third evaluation in automatic speech recognition and associated technologies in a series which started with ESTER. This evaluation proposed some new challenges, by proposing TV and radio shows with prepared and spontaneous speech, annotation and evaluation of overlapping speech, a cross-show condition in speaker diarization, and new, complex but very informative named entities in the information extraction task. This paper presents the whole campaign, including the data annotated, the metrics used and the anonymized system results. All the data created in the evaluation, hopefully including system outputs, will be distributed through the ELRA catalogue in the future.
Code (0)
등록된 구현이 없습니다.
Tasks
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speaker-diarizationSpeaker DiarizationSpeaker Recognitionspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
The ETAPE corpus for the evaluation of speech-based TV content processing in the French language
The paper presents a comprehensive overview of existing data for the evaluation of spoken content processing in a multimedia framework for the French language. We focus on the ETAPE corpus which will be made publicly ava…
Speech RecognitionWhere are we in Named Entity Recognition from Speech?
Named entity recognition (NER) from speech is usually made through a pipeline process that consists in (i) processing audio using an automatic speech recognition system (ASR) and (ii) applying a NER to the ASR outputs. T…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Entity Extraction using GANnamed-entity-recognition+5MetaPerturb: Transferable Regularizer for Heterogeneous Tasks and Architectures
Regularization and transfer learning are two popular techniques to enhance generalization on unseen data, which is a fundamental problem of machine learning. Regularization techniques are versatile, as they are task- and…
Meta-LearningTransfer LearningOverlap-aware diarization: resegmentation using neural end-to-end overlapped speech detection
We address the problem of effectively handling overlapping speech in a diarization system. First, we detail a neural Long Short-Term Memory-based architecture for overlap detection. Secondly, detected overlap regions are…
O\`u en sommes-nous dans la reconnaissance des entit\'es nomm\'ees structur\'ees \`a partir de la parole ? (Where are we in Named Entity Recognition from speech ?)
La reconnaissance des entit{\'e}s nomm{\'e}es (REN) {\`a} partir de la parole est traditionnellement effectu{\'e}e par l{'}interm{\'e}diaire d{'}une cha{\^\i}ne de composants, exploitant un syst{\`e}me de reconnaissance …
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)