paper-with-me

홈 › Papers

Best of Both Worlds: Robust Accented Speech Recognition with Adversarial Transfer Learning

2021-03-10 · Nilaksh Das, Sravan Bodapati, Monica Sunkara, Sundararajan Srinivasan, Duen Horng Chau

Training deep neural networks for automatic speech recognition (ASR) requires large amounts of transcribed speech. This becomes a bottleneck for training robust models for accented speech which typically contains high variability in pronunciation and other semantics, since obtaining large amounts of annotated accented data is both tedious and costly. Often, we only have access to large amounts of unannotated speech from different accents. In this work, we leverage this unannotated data to provide semantic regularization to an ASR model that has been trained only on one accent, to improve its performance for multiple accents. We propose Accent Pre-Training (Acc-PT), a semi-supervised training strategy that combines transfer learning and adversarial training. Our approach improves the performance of a state-of-the-art ASR model by 33% on average over the baseline across multiple accents, training only on annotated samples from one standard accent, and as little as 105 minutes of unannotated speech from a target accent.

📄 PDF Abstract BibTeX arXiv:2103.05834

Code (0)

등록된 구현이 없습니다.

Tasks

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis

2024-07-04 · Cong-Thanh Do, Shuhei Imai, Rama Doddipatla, Thomas Hain

This paper investigates the use of unsupervised text-to-speech synthesis (TTS) as a data augmentation method to improve accented speech recognition. TTS systems are trained with a small amount of accented speech training…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentation+7

Accented Speech Recognition: A Survey

2021-04-21 · Arthur Hinsvark, Natalie Delworth, Miguel Del Rio, Quinten McNamara 외

Automatic Speech Recognition (ASR) systems generalize poorly on accented speech. The phonetic and linguistic variability of accents present hard challenges for ASR systems today in both data collection and modeling strat…

Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Feature Engineering+3

Exploring data augmentation in bias mitigation against non-native-accented speech

2023-12-24 · Yuanyuan Zhang, Aaricia Herygers, Tanvina Patel, Zhengjun Yue 외

Automatic speech recognition (ASR) should serve every speaker, not only the majority ``standard'' speakers of a language. In order to build inclusive ASR, mitigating the bias against speaker groups who speak in a ``non-s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2

Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech

2021-09-14 · EMNLP 2021 11 · Katrin Tomanek, Vicky Zayats, Dirk Padfield, Kara Vaillancourt 외

Automatic Speech Recognition (ASR) systems are often optimized to work best for speakers with canonical speech patterns. Unfortunately, these systems perform poorly when tested on atypical speech and heavily accented spe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Multi-pass Training and Cross-information Fusion for Low-resource End-to-end Accented Speech Recognition

2023-06-20 · Xuefei Wang, Yanhua Long, Yijie Li, Haoran Wei

Low-resource accented speech recognition is one of the important challenges faced by current ASR technology in practical applications. In this study, we propose a Conformer-based architecture, called Aformer, to leverage…

Accented Speech Recognitionspeech-recognitionSpeech Recognition