paper-with-me

홈 › Papers

Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English with Transfer Learning

2021-10-01 · Toshiko Shibano, Xinyi Zhang, Mia Taige Li, Haejin Cho, Peter Sullivan, Muhammad Abdul-Mageed

To address the performance gap of English ASR models on L2 English speakers, we evaluate fine-tuning of pretrained wav2vec 2.0 models (Baevski et al., 2020; Xu et al., 2021) on L2-ARCTIC, a non-native English speech corpus (Zhao et al., 2018) under different training settings. We compare \textbf{(a)} models trained with a combination of diverse accents to ones trained with only specific accents and \textbf{(b)} results from different single-accent models. Our experiments demonstrate the promise of developing ASR models for non-native English speakers, even with small amounts of L2 training data and even without a language model. Our models also excel in the zero-shot setting where we train on multiple L2 datasets and test on a blind L2 test set.

📄 PDF Abstract BibTeX arXiv:2110.00678

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modellingspeech-recognitionSpeech RecognitionTransfer Learning

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English

2021-11-01 · ICNLSP 2021 11 · Toshiko Shibano, Xinyi Zhang, Mia Taige Li, Haejin Cho 외
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

MUSCAT: MUltilingual, SCientific ConversATion Benchmark

2026-04-17 · Supriti Sinhamahapatra, Thai-Binh Nguyen, Yiğit Oğuz, Enes Ugan 외 arxiv

The goal of multilingual speech technology is to facilitate seamless communication between individuals speaking different languages, creating the experience as though everyone were a multilingual speaker. To create this …

Speaker DiarizationSpeech Recognition

Hey ASR System! Why Aren't You More Inclusive? Automatic Speech Recognition Systems' Bias and Proposed Bias Mitigation Techniques. A Literature Review

2022-11-17 · Mikel K. Ngueajio, Gloria Washington

Speech is the fundamental means of communication between humans. The advent of AI and sophisticated speech technologies have led to the rapid proliferation of human-to-computer-based interactions, fueled primarily by Aut…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch

2024-12-20 · Xingchen Song, Chengdong Liang, BinBin Zhang, Pengshen Zhang 외

Large Automatic Speech Recognition (ASR) models demand a vast number of parameters, copious amounts of data, and significant computational resources during the training process. However, such models can merely be deploye…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

運用Python結合語音辨識及合成技術於自動化音文同步之實作(A Python Implementation of Automatic Speech-text Synchronization Using Speech Recognition and Text-to-Speech Technology)[In Chinese]

2015-10-01 · ROCLINGIJCLCLP 2015 10 · ChunHan Lai, Chao-Kai Chang, Ren-Yuan Lyu
speech-recognitionSpeech Recognitiontext-to-speechText to Speech