paper-with-me

홈 › Papers

Foreign English Accent Adjustment by Learning Phonetic Patterns

2018-07-09 · Fedor Kitashov, Elizaveta Svitanko, Debojyoti Dutta

State-of-the-art automatic speech recognition (ASR) systems struggle with the lack of data for rare accents. For sufficiently large datasets, neural engines tend to outshine statistical models in most natural language processing problems. However, a speech accent remains a challenge for both approaches. Phonologists manually create general rules describing a speaker's accent, but their results remain underutilized. In this paper, we propose a model that automatically retrieves phonological generalizations from a small dataset. This method leverages the difference in pronunciation between a particular dialect and General American English (GAE) and creates new accented samples of words. The proposed model is able to learn all generalizations that previously were manually obtained by phonologists. We use this statistical method to generate a million phonological variations of words from the CMU Pronouncing Dictionary and train a sequence-to-sequence RNN to recognize accented words with 59% accuracy.

📄 PDF Abstract BibTeX arXiv:1807.03625

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Improved Accent Classification Combining Phonetic Vowels with Acoustic Features

2016-02-24 · Zhenhao Ge

Researches have shown accent classification can be improved by integrating semantic information into pure acoustic approach. In this work, we combine phonetic knowledge, such as vowels, with enhanced acoustic features to…

ClassificationGeneral Classification

Accented Text-to-Speech Synthesis with Limited Data

2023-05-08 · Xuehao Zhou, Mingyang Zhang, Yi Zhou, Zhizheng Wu 외

This paper presents an accented text-to-speech (TTS) synthesis framework with limited training data. We study two aspects concerning accent rendering: phonetic (phoneme difference) and prosodic (pitch pattern and phoneme…

Speech Synthesistext-to-speechText to SpeechText-To-Speech Synthesis

FROST-EMA: Finnish and Russian Oral Speech Dataset of Electromagnetic Articulography Measurements with L1, L2 and Imitated L2 Accents

2025-06-10 · Satu Hopponen, Tomi Kinnunen, Alexandre Nikolaev, Rosa González Hautamäki 외

We introduce a new FROST-EMA (Finnish and Russian Oral Speech Dataset of Electromagnetic Articulography) corpus. It consists of 18 bilingual speakers, who produced speech in their native language (L1), second language (L…

Speaker Verification

Synthetic Cross-accent Data Augmentation for Automatic Speech Recognition

2023-03-01 · Philipp Klumpp, Pooja Chitkara, Leda Sari, Prashant Serai 외

The awareness for biased ASR datasets or models has increased notably in recent years. Even for English, despite a vast amount of available training data, systems perform worse for non-native speakers. In this work, we i…

Automatic Speech RecognitionData Augmentationspeech-recognitionSpeech Recognition

Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives

2025-05-21 · Gowtham Premananth, Vinith Kugathasan, Carol Espy-Wilson

Advancements in AI-driven speech-based applications have transformed diverse industries ranging from healthcare to customer service. However, the increasing prevalence of non-native accented speech in global interactions…