Accented Speech Recognition
4개 벤치마크 · 논문 22편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
Deep Speech: Scaling up end-to-end speech recognition
Accented Speech Recognition With Accent-specific Codebooks
Advancing African-Accented Speech Recognition: Epistemic Uncertainty-Driven Data Selection for Generalizable ASR Models
Goodness of Pronunciation Pipelines for OOV Problem
Papers
ROMPAR: Morphological Completion and Demographic Unlearning for Romanian-Accented Speech Recognition
Automated transcription of parliamentary proceedings faces significant hurdles due to demographic bias, dialectal variation, and technical artifacts such as utterance truncation during segmentation. This paper introduces…
Accented Speech RecognitionMixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
Accented speech remains a persistent challenge for automatic speech recognition (ASR), as most models are trained on data dominated by a few high-resource English varieties, leading to substantial performance degradation…
Accented Speech RecognitionLeveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative Analysis
Large-scale training corpora have significantly improved the performance of ASR models. Unfortunately, due to the relative scarcity of data, Chinese accents and dialects remain a challenge for most ASR models. Recent adv…
Accented Speech RecognitionSelf-Supervised Learningspeech-recognitionSpeech RecognitionGE2E-AC: Generalized End-to-End Loss Training for Accent Classification
Accent classification or AC is a task to predict the accent type of an input utterance, and it can be used as a preliminary step toward accented speech recognition and accent conversion. Existing studies have often achie…
Accented Speech RecognitionClassificationspeech-recognitionSpeech RecognitionImproving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis
This paper investigates the use of unsupervised text-to-speech synthesis (TTS) as a data augmentation method to improve accented speech recognition. TTS systems are trained with a small amount of accented speech training…
Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentation+7Accented Speech Recognition With Accent-specific Codebooks
Speech accents pose a significant challenge to state-of-the-art automatic speech recognition (ASR) systems. Degradation in performance across underrepresented accents is a severe deterrent to the inclusive adoption of AS…
Accented Speech RecognitionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognition+1