paper-with-me

Papers

Morpho-Syntactic Study of Errors from Speech Recognition System

2014-05-01 · LREC 2014 5 · Maria Goryainova, Cyril Grouin, Sophie Rosset, Ioana Vasilescu

The study provides an original standpoint of the speech transcription errors by focusing on the morpho-syntactic features of the erroneous chunks and of the surrounding left and right context. The typology concerns the forms, the lemmas and the POS involved in erroneous chunks, and in the surrounding contexts. Comparison with error free contexts are also provided. The study is conducted on French. Morpho-syntactic analysis underlines that three main classes are particularly represented in the erroneous chunks: (i) grammatical words (to, of, the), (ii) auxiliary verbs (has, is), and (iii) modal verbs (should, must). Such items are widely encountered in the ASR outputs as frequent candidates to transcription errors. The analysis of the context points out that some left 3-grams contexts (e.g., repetitions, that is disfluencies, bracketing formulas such as {``}cÂ’est{''}, etc.) may be better predictors than others. Finally, the surface analysis conducted through a Levensthein distance analysis, highlighted that the most common distance is of 2 characters and mainly involves differences between inflected forms of a unique item.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Named Entity Recognition (NER)POSQuestion Answeringspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Automatic Speech Recognition Biases in Newcastle English: an Error Analysis

2025-06-19 · Dana Serditova, Kevin Tang, Jochen Steffens

Automatic Speech Recognition (ASR) systems struggle with regional dialects due to biased training which favours mainstream varieties. While previous research has identified racial, age, and gender biases in ASR, regional…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition

2026-04-30 · Thibault Bañeras-Roux, Mickaël Rouvier, Jane Wottawa, Richard Dufour arxiv

Evaluating automatic speech recognition (ASR) systems is a classical but difficult and still open problem, which often boils down to focusing only on the word error rate (WER). However, this metric suffers from many limi…

Speech Recognition

Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information

2024-04-23 · Chihiro Taguchi, Jefferson Saransig, Dayana Velásquez, David Chiang

This paper presents Killkan, the first dataset for automatic speech recognition (ASR) in the Kichwa language, an indigenous language of Ecuador. Kichwa is an extremely low-resource endangered language, and there have bee…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Enhanced CORILGA: Introducing the Automatic Phonetic Alignment Tool for Continuous Speech

2016-05-01 · LREC 2016 5 · Roberto Seara, Marta Martinez, Roc{\'\i}o Varela, Carmen Garc{\'\i}a Mateo 외

The {``}Corpus Oral Informatizado da Lingua Galega (CORILGA){''} project aims at building a corpus of oral language for Galician, primarily designed to study the linguistic variation and change. This project is currently…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Sentencespeech-recognition+1

Understanding the effects of word-level linguistic annotations in under-resourced neural machine translation

2020-12-01 · COLING 2020 8 · V{\'\i}ctor M. S{\'a}nchez-Cartagena, Juan Antonio P{\'e}rez-Ortiz, Felipe S{\'a}nchez-Mart{\'\i}nez

This paper studies the effects of word-level linguistic annotations in under-resourced neural machine translation, for which there is incomplete evidence in the literature. The study covers eight language pairs, differen…

Machine TranslationTAGTranslation