paper-with-me

Papers

Evaluating Automatic Speech Recognition Systems in Comparison With Human Perception Results Using Distinctive Feature Measures

2016-12-13 · Xiang Kong, Jeung-Yoon Choi, Stefanie Shattuck-Hufnagel

This paper describes methods for evaluating automatic speech recognition (ASR) systems in comparison with human perception results, using measures derived from linguistic distinctive features. Error patterns in terms of manner, place and voicing are presented, along with an examination of confusion matrices via a distinctive-feature-distance metric. These evaluation methods contrast with conventional performance criteria that focus on the phone or word level, and are intended to provide a more detailed profile of ASR system performance,as well as a means for direct comparison with human perception results at the sub-phonemic level.

📄 PDF Abstract BibTeX arXiv:1612.03990

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

ESB: A Benchmark For Multi-Domain End-to-End Speech Recognition

2022-10-24 · Sanchit Gandhi, Patrick von Platen, Alexander M. Rush

Speech recognition applications cover a range of different audio and text distributions, with different speaking styles, background noise, transcription punctuation and character casing. However, many speech recognition …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)BenchmarkingSpeech Recognition

NoRefER: a Referenceless Quality Metric for Automatic Speech Recognition via Semi-Supervised Language Model Fine-Tuning with Contrastive Learning

2023-06-21 · Kamer Ali Yuksel, Thiago Ferreira, Golara Javadi, Mohamed El-Badrashiny 외

This paper introduces NoRefER, a novel referenceless quality metric for automatic speech recognition (ASR) systems. Traditional reference-based metrics for evaluating ASR systems require costly ground-truth transcripts. …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Contrastive LearningLanguage Modeling+4

CEASR: A Corpus for Evaluating Automatic Speech Recognition

2020-05-01 · LREC 2020 5 · Malgorzata Anna Ulasik, Manuela H{\"u}rlimann, Fabian Germann, Esin Gedik 외

In this paper, we present CEASR, a Corpus for Evaluating the quality of Automatic Speech Recognition (ASR). It is a data set based on public speech corpora, containing metadata along with transcripts generated by several…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Framework for Curating Speech Datasets and Evaluating ASR Systems: A Case Study for Polish

2024-07-18 · Michał Junczyk

Speech datasets available in the public domain are often underutilized because of challenges in discoverability and interoperability. A comprehensive framework has been designed to survey, catalog, and curate available s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition

2024-07-31 · Nick Rossenbach, Ralf Schlüter, Sakriani Sakti

The rapid development of neural text-to-speech (TTS) systems enabled its usage in other areas of natural language processing such as automatic speech recognition (ASR) or spoken language translation (SLT). Due to the lar…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderModel Selection+5