paper-with-me

홈 › Papers

Modelling low-resource accents without accent-specific TTS frontend

2023-01-11 · Georgi Tinchev, Marta Czarnowska, Kamil Deja, Kayoko Yanagisawa, Marius Cotescu

This work focuses on modelling a speaker's accent that does not have a dedicated text-to-speech (TTS) frontend, including a grapheme-to-phoneme (G2P) module. Prior work on modelling accents assumes a phonetic transcription is available for the target accent, which might not be the case for low-resource, regional accents. In our work, we propose an approach whereby we first augment the target accent data to sound like the donor voice via voice conversion, then train a multi-speaker multi-accent TTS model on the combination of recordings and synthetic data, to generate the donor's voice speaking in the target accent. Throughout the procedure, we use a TTS frontend developed for the same language but a different accent. We show qualitative and quantitative analysis where the proposed strategy achieves state-of-the-art results compared to other generative models. Our work demonstrates that low resource accents can be modelled with relatively little data and without developing an accent-specific TTS frontend. Audio samples of our model converting to multiple accents are available on our web page.

📄 PDF Abstract BibTeX arXiv:2301.04606

Code (0)

등록된 구현이 없습니다.

Tasks

text-to-speechText to SpeechVoice Conversion

Similar Papers 제목 키워드 기반

Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition

2022-01-24 · Xurong Xie, Xiang Sui, Xunying Liu, Lan Wang

The Mandarin Chinese language is known to be strongly influenced by a rich set of regional accents, while Mandarin speech with each accent is quite low resource. Hence, an important task in Mandarin speech recognition is…

Acoustic Modellingspeech-recognitionSpeech Recognition

Effects of Speaker Count, Duration, and Accent Diversity on Zero-Shot Accent Robustness in Low-Resource ASR

2025-06-04 · Zheng-Xin Yong, Vineel Pratap, Michael Auli, Jean Maillard

To build an automatic speech recognition (ASR) system that can serve everyone in the world, the ASR needs to be robust to a wide range of accents including unseen accents. We systematically study how three different vari…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Low-resource Accent Classification in Geographically-proximate Settings: A Forensic and Sociophonetics Perspective

2022-06-26 · Qingcheng Zeng, Dading Chong, Peilin Zhou, Jie Yang

Accented speech recognition and accent classification are relatively under-explored research areas in speech technology. Recently, deep learning-based methods and Transformer-based pretrained models have achieved superb …

Accented Speech RecognitionClassificationspeech-recognitionSpeech Recognition

Open-source Multi-speaker Corpora of the English Accents in the British Isles

2020-05-01 · LREC 2020 5 · Isin Demirsahin, Oddur Kjartansson, Alex Gutkin, er 외

This paper presents a dataset of transcribed high-quality audio of English sentences recorded by volunteers speaking with different accents of the British Isles. The dataset is intended for linguistic analysis as well as…

Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition

2026-02-02 · Wonjun Lee, Hyounghun Kim, Gary Geunbae Lee arxiv

Accented speech remains a persistent challenge for automatic speech recognition (ASR), as most models are trained on data dominated by a few high-resource English varieties, leading to substantial performance degradation…

Accented Speech Recognition