paper-with-me

홈 › Papers

Evaluation of Off-the-shelf Speech Recognizers on Different Accents in a Dialogue Domain

2022-06-01 · LREC 2022 6 · Divya Tadimeti, Kallirroi Georgila, David Traum

We evaluate several publicly available off-the-shelf (commercial and research) automatic speech recognition (ASR) systems on dialogue agent-directed English speech from speakers with General American vs. non-American accents. Our results show that the performance of the ASR systems for non-American accents is considerably worse than for General American accents. Depending on the recognizer, the absolute difference in performance between General American accents and all non-American accents combined can vary approximately from 2% to 12%, with relative differences varying approximately between 16% and 49%. This drop in performance becomes even larger when we consider specific categories of non-American accents indicating a need for more diligent collection of and training on non-native English speaker data in order to narrow this performance gap. There are performance differences across ASR systems, and while the same general pattern holds, with more errors for non-American accents, there are some accents for which the best recognizer is different than in the overall case. We expect these results to be useful for dialogue system designers in developing more robust inclusive dialogue systems, and for ASR providers in taking into account performance requirements for different accents.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Evaluation of Off-the-shelf Speech Recognizers Across Diverse Dialogue Domains

2020-05-01 · LREC 2020 5 · Kallirroi Georgila, Anton Leuski, Volodymyr Yanov, David Traum

We evaluate several publicly available off-the-shelf (commercial and research) automatic speech recognition (ASR) systems across diverse dialogue domains (in US-English). Our evaluation is aimed at non-experts with limit…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Benchmarking_Fast_Domain_Adaptation_for_Unsupervised_Speech_Units

2026-08-27 · Robin San Roman, Manel Khentout, Tu Anh Nguyen, Paul Michel 외 arxiv

Representation learning has attracted great atten- tion and managed to reach good performances as a pretraining method for downstream tasks or as a first step towards unsu- pervised speech modeling. Yet, little is known …

Representation Learning

AccentFold: A Journey through African Accents for Zero-Shot ASR Adaptation to Target Accents

2024-02-02 · Abraham Toluwase Owodunni, Aditya Yadavalli, Chris Chinenye Emezue, Tobi Olatunji 외

Despite advancements in speech recognition, accented speech remains challenging. While previous approaches have focused on modeling techniques or creating accented speech datasets, gathering sufficient data for the multi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

An Empirical Study on L2 Accents of Cross-lingual Text-to-Speech Systems via Vowel Space

2022-11-06 · JIhwan Lee, Jae-Sung Bae, Seongkyu Mun, Heejin Choi 외

With the recent developments in cross-lingual Text-to-Speech (TTS) systems, L2 (second-language, or foreign) accent problems arise. Moreover, running a subjective evaluation for such cross-lingual TTS systems is troubles…

text-to-speechText to Speech

How Accents Confound: Probing for Accent Information in End-to-End Speech Recognition Systems

2020-07-01 · ACL 2020 6 · Archiki Prasad, Preethi Jyothi

In this work, we present a detailed analysis of how accent information is reflected in the internal representation of speech in an end-to-end automatic speech recognition (ASR) system. We use a state-of-the-art end-to-en…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition