paper-with-me

홈 › Papers

Automatic Speech Recognition of African American English: Lexical and Contextual Effects

2025-06-07 · Hamid Mojarad, Kevin Tang

Automatic Speech Recognition (ASR) models often struggle with the phonetic, phonological, and morphosyntactic features found in African American English (AAE). This study focuses on two key AAE variables: Consonant Cluster Reduction (CCR) and ING-reduction. It examines whether the presence of CCR and ING-reduction increases ASR misrecognition. Subsequently, it investigates whether end-to-end ASR systems without an external Language Model (LM) are more influenced by lexical neighborhood effect and less by contextual predictability compared to systems with an LM. The Corpus of Regional African American Language (CORAAL) was transcribed using wav2vec 2.0 with and without an LM. CCR and ING-reduction were detected using the Montreal Forced Aligner (MFA) with pronunciation expansion. The analysis reveals a small but significant effect of CCR and ING on Word Error Rate (WER) and indicates a stronger presence of lexical neighborhood effect in ASR systems without LMs.

📄 PDF Abstract BibTeX arXiv:2506.06888

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modellingspeech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Dialect-Specific Models for Automatic Speech Recognition of African American Vernacular English

2019-09-01 · RANLP 2019 9 · Rachel Dorn

African American Vernacular English (AAVE) is a widely-spoken dialect of English, yet it is under-represented in major speech corpora. As a result, speakers of this dialect are often misunderstood by NLP applications. Th…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Improving Speech Recognition for African American English With Audio Classification

2023-09-16 · Shefali Garg, Zhouyuan Huo, Khe Chai Sim, Suzan Schwartz 외

Automatic speech recognition (ASR) systems have been shown to have large quality disparities between the language varieties they are intended or expected to recognize. One way to mitigate this is to train or fine-tune mo…

Audio ClassificationAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Form+2

Self-supervised Speech Representations Still Struggle with African American Vernacular English

2024-08-26 · Kalvin Chang, Yi-Hui Chou, Jiatong Shi, Hsuan-Ming Chen 외

Underperformance of ASR systems for speakers of African American Vernacular English (AAVE) and other marginalized language varieties is a well-documented phenomenon, and one that reinforces the stigmatization of these va…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Self-Supervised Learningspeech-recognition+1

Quantification of stylistic differences in human- and ASR-produced transcripts of African American English

2024-09-04 · Annika Heuser, Tyler Kendall, Miguel Del Rio, Quinten McNamara 외

Common measures of accuracy used to assess the performance of automatic speech recognition (ASR) systems, as well as human transcribers, conflate multiple sources of error. Stylistic differences, such as verbatim vs non-…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Layer-wise Probing of wav2vec 2.0 and Whisper for Consonant Cluster Reduction in African American English

2026-06-22 · Hamid Mojarad, Kevin Tang arxiv

Self-supervised and supervised speech models are increasingly used to investigate which linguistic information their internal representations encode, and at what level of abstraction they encode it. One underexplored phe…

Speech Recognition