paper-with-me

Papers

LID Models are Actually Accent Classifiers: Implications and Solutions for LID on Accented Speech

2025-05-31 · Niyati Bafna, Matthew Wiesner

Prior research indicates that LID model performance significantly declines on accented speech; however, the specific causes, extent, and characterization of these errors remain under-explored. (i) We identify a common failure mode on accented speech whereby LID systems often misclassify L2 accented speech as the speaker's native language or a related language. (ii) We present evidence suggesting that state-of-the-art models are invariant to permutations of short spans of speech, implying they classify on the basis of short phonotactic features indicative of accent rather than language. Our analysis reveals a simple method to enhance model robustness to accents through input chunking. (iii) We present an approach that integrates sequence-level information into our model without relying on monolingual ASR systems; this reduces accent-language confusion and significantly enhances performance on accented speech while maintaining comparable results on standard LID.

📄 PDF Abstract BibTeX arXiv:2506.00628

Code (0)

등록된 구현이 없습니다.

Tasks

Chunking

Similar Papers 제목 키워드 기반

Accented Text-to-Speech Synthesis with Limited Data

2023-05-08 · Xuehao Zhou, Mingyang Zhang, Yi Zhou, Zhizheng Wu 외

This paper presents an accented text-to-speech (TTS) synthesis framework with limited training data. We study two aspects concerning accent rendering: phonetic (phoneme difference) and prosodic (pitch pattern and phoneme…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis

Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis

2024-07-04 · Cong-Thanh Do, Shuhei Imai, Rama Doddipatla, Thomas Hain

This paper investigates the use of unsupervised text-to-speech synthesis (TTS) as a data augmentation method to improve accented speech recognition. TTS systems are trained with a small amount of accented speech training…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentation+7

Learning-free L2-Accented Speech Generation using Phonological Rules

2026-03-08 · Thanathai Lertpetchpun, Yoonjeong Lee, Jihwan Lee, Tiantian Feng 외 arxiv

Accent plays a crucial role in speaker identity and inclusivity in speech technologies. Existing accented text-to-speech (TTS) systems either require large-scale accented datasets or lack fine-grained phoneme-level contr…

Emirati-Accented Speaker Identification in Stressful Talking Conditions

2019-09-28 · Ismail Shahin, Ali Bou Nassif

This research is dedicated to improving text-independent Emirati-accented speaker identification performance in stressful talking conditions using three distinct classifiers: First-Order Hidden Markov Models (HMM1s), Sec…

Speaker Identification

Domain Adversarial Training for Accented Speech Recognition

2018-06-07 · Sining Sun, Ching-Feng Yeh, Mei-Yuh Hwang, Mari Ostendorf 외

In this paper, we propose a domain adversarial training (DAT) algorithm to alleviate the accented speech recognition problem. In order to reduce the mismatch between labeled source domain data ("standard" accent) and unl…

Accented Speech RecognitionMulti-Task Learningspeech-recognitionSpeech Recognition