paper-with-me

Papers

Accented Speech Recognition: Benchmarking, Pre-training, and Diverse Data

2022-05-16 · Alëna Aksënova, Zhehuai Chen, Chung-Cheng Chiu, Daan van Esch, Pavel Golik, Wei Han, Levi King, Bhuvana Ramabhadran, Andrew Rosenberg, Suzan Schwartz, Gary Wang

Building inclusive speech recognition systems is a crucial step towards developing technologies that speakers of all language varieties can use. Therefore, ASR systems must work for everybody independently of the way they speak. To accomplish this goal, there should be available data sets representing language varieties, and also an understanding of model configuration that is the most helpful in achieving robust understanding of all types of speech. However, there are not enough data sets for accented speech, and for the ones that are already available, more training approaches need to be explored to improve the quality of accented speech recognition. In this paper, we discuss recent progress towards developing more inclusive ASR systems, namely, the importance of building new data sets representing linguistic diversity, and exploring novel training approaches to improve performance for all users. We address recent directions within benchmarking ASR systems for accented speech, measure the effects of wav2vec 2.0 pre-training on accented speech recognition, and highlight corpora relevant for diverse ASR evaluations.

📄 PDF Abstract BibTeX arXiv:2205.08014

Code (0)

등록된 구현이 없습니다.

Tasks

Accented Speech RecognitionBenchmarkingDiversityspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Coupled Training of Sequence-to-Sequence Models for Accented Speech Recognition

2020-05-14 · Vinit Unni, Nitish Joshi, Preethi Jyothi

Accented speech poses significant challenges for state-of-the-art automatic speech recognition (ASR) systems. Accent is a property of speech that lasts throughout an utterance in varying degrees of strength. This makes i…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Decoder+2

Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis

2024-07-04 · Cong-Thanh Do, Shuhei Imai, Rama Doddipatla, Thomas Hain

This paper investigates the use of unsupervised text-to-speech synthesis (TTS) as a data augmentation method to improve accented speech recognition. TTS systems are trained with a small amount of accented speech training…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentation+7

Domain Adversarial Training for Accented Speech Recognition

2018-06-07 · Sining Sun, Ching-Feng Yeh, Mei-Yuh Hwang, Mari Ostendorf 외

In this paper, we propose a domain adversarial training (DAT) algorithm to alleviate the accented speech recognition problem. In order to reduce the mismatch between labeled source domain data ("standard" accent) and unl…

Accented Speech RecognitionMulti-Task Learningspeech-recognitionSpeech Recognition

Exploring data augmentation in bias mitigation against non-native-accented speech

2023-12-24 · Yuanyuan Zhang, Aaricia Herygers, Tanvina Patel, Zhengjun Yue 외

Automatic speech recognition (ASR) should serve every speaker, not only the majority ``standard'' speakers of a language. In order to build inclusive ASR, mitigating the bias against speaker groups who speak in a ``non-s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition

2024-08-05

Modern automatic speech recognition (ASR) systems are typically trained on more than tens of thousands hours of speech data, which is one of the main factors for their great success. However, the distribution of such dat…