paper-with-me

홈 › Papers

English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System

2021-05-09 · Guillermo Cámbara, Alex Peiró-Lilja, Mireia Farrús, Jordi Luque

Nowadays, research in speech technologies has gotten a lot out thanks to recently created public domain corpora that contain thousands of recording hours. These large amounts of data are very helpful for training the new complex models based on deep learning technologies. However, the lack of dialectal diversity in a corpus is known to cause performance biases in speech systems, mainly for underrepresented dialects. In this work, we propose to evaluate a state-of-the-art automatic speech recognition (ASR) deep learning-based model, using unseen data from a corpus with a wide variety of labeled English accents from different countries around the world. The model has been trained with 44.5K hours of English speech from an open access corpus called Multilingual LibriSpeech, showing remarkable results in popular benchmarks. We test the accuracy of such ASR against samples extracted from another public corpus that is continuously growing, the Common Voice dataset. Then, we present graphically the accuracy in terms of Word Error Rate of each of the different English included accents, showing that there is indeed an accuracy bias in terms of accentual variety, favoring the accents most prevalent in the training corpus.

📄 PDF Abstract BibTeX arXiv:2105.05041

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Foreign English Accent Adjustment by Learning Phonetic Patterns

2018-07-09 · Fedor Kitashov, Elizaveta Svitanko, Debojyoti Dutta

State-of-the-art automatic speech recognition (ASR) systems struggle with the lack of data for rare accents. For sufficiently large datasets, neural engines tend to outshine statistical models in most natural language pr…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

How Accents Confound: Probing for Accent Information in End-to-End Speech Recognition Systems

2020-07-01 · ACL 2020 6 · Archiki Prasad, Preethi Jyothi

In this work, we present a detailed analysis of how accent information is reflected in the internal representation of speech in an end-to-end automatic speech recognition (ASR) system. We use a state-of-the-art end-to-en…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

AccentDB: A Database of Non-Native English Accents to Assist Neural Speech Recognition

2020-05-16 · LREC 2020 5 · Afroz Ahamad, Ankit Anand, Pranesh Bhargava

Modern Automatic Speech Recognition (ASR) technology has evolved to identify the speech spoken by native speakers of a language very well. However, identification of the speech spoken by non-native speakers continues to …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Multilingual Approach to Joint Speech and Accent Recognition with DNN-HMM Framework

2020-10-22 · Yizhou Peng, Jicheng Zhang, Haobo Zhang, HaiHua Xu 외

Human can recognize speech, as well as the peculiar accent of the speech simultaneously. However, present state-of-the-art ASR system can rarely do that. In this paper, we propose a multilingual approach to recognizing E…

speech-recognitionSpeech RecognitionTransfer Learning

CommonAccent: Exploring Large Acoustic Pretrained Models for Accent Classification Based on Common Voice

2023-05-29 · Juan Zuluaga-Gomez, Sara Ahmed, Danielius Visockas, Cem Subakan

Despite the recent advancements in Automatic Speech Recognition (ASR), the recognition of accented speech still remains a dominant problem. In order to create more inclusive ASR systems, research has shown that the integ…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Classificationspeech-recognition+1