paper-with-me

홈 › Papers

How Accents Confound: Probing for Accent Information in End-to-End Speech Recognition Systems

2020-07-01 · ACL 2020 6 · Archiki Prasad, Preethi Jyothi

In this work, we present a detailed analysis of how accent information is reflected in the internal representation of speech in an end-to-end automatic speech recognition (ASR) system. We use a state-of-the-art end-to-end ASR system, comprising convolutional and recurrent layers, that is trained on a large amount of US-accented English speech and evaluate the model on speech samples from seven different English accents. We examine the effects of accent on the internal representation using three main probing techniques: a) Gradient-based explanation methods, b) Information-theoretic measures, and c) Outputs of accent and phone classifiers. We find different accents exhibiting similar trends irrespective of the probing technique used. We also find that most accent information is encoded within the first recurrent layer, which is suggestive of how one could adapt such an end-to-end model to learn representations that are invariant to accents.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents

2024-02-02 · Abraham Toluwase Owodunni, Aditya Yadavalli, Chris Chinenye Emezue, Tobi Olatunji 외

Despite advancements in speech recognition, accented speech remains challenging. While previous approaches have focused on modeling techniques or creating accented speech datasets, gathering sufficient data for the multi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Accented Speech Recognition With Accent-specific Codebooks

2023-10-24 · Darshan Prabhu, Preethi Jyothi, Sriram Ganapathy, Vinit Unni

Speech accents pose a significant challenge to state-of-the-art automatic speech recognition (ASR) systems. Degradation in performance across underrepresented accents is a severe deterrent to the inclusive adoption of AS…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1

DITTO: Data-efficient and Fair Targeted Subset Selection for ASR Accent Adaptation

2021-10-10 · Suraj Kothawade, Anmol Mekala, Chandra Sekhara D, Mayank Kothyari 외

State-of-the-art Automatic Speech Recognition (ASR) systems are known to exhibit disparate performance on varying speech accents. To improve performance on a specific target accent, a commonly adopted solution is to fine…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition

2018-02-07 · Xuesong Yang, Kartik Audhkhasi, Andrew Rosenberg, Samuel Thomas 외

The performance of automatic speech recognition systems degrades with increasing mismatch between the training and testing scenarios. Differences in speaker accents are a significant source of such mismatch. The traditio…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Evaluation of Off-the-shelf Speech Recognizers on Different Accents in a Dialogue Domain

2022-06-01 · LREC 2022 6 · Divya Tadimeti, Kallirroi Georgila, David Traum

We evaluate several publicly available off-the-shelf (commercial and research) automatic speech recognition (ASR) systems on dialogue agent-directed English speech from speakers with General American vs. non-American acc…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition