paper-with-me

Papers

Exploring Generative Error Correction for Dysarthric Speech Recognition

2025-05-26 · Moreno La Quatra, Alkis Koudounas, Valerio Mario Salerno, Sabato Marco Siniscalchi

Despite the remarkable progress in end-to-end Automatic Speech Recognition (ASR) engines, accurately transcribing dysarthric speech remains a major challenge. In this work, we proposed a two-stage framework for the Speech Accessibility Project Challenge at INTERSPEECH 2025, which combines cutting-edge speech recognition models with LLM-based generative error correction (GER). We assess different configurations of model scales and training strategies, incorporating specific hypothesis selection to improve transcription accuracy. Experiments on the Speech Accessibility Project dataset demonstrate the strength of our approach on structured and spontaneous speech, while highlighting challenges in single-word recognition. Through comprehensive analysis, we provide insights into the complementary roles of acoustic and linguistic modeling in dysarthric speech recognition

📄 PDF Abstract BibTeX arXiv:2505.20163

Code (1)

morenolaquatra/ger4dys 공식 구현

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Arabic Dysarthric Speech Recognition Using Adversarial and Signal-Based Augmentation

2023-06-07 · Massa Baali, Ibrahim Almakky, Shady Shehata, Fakhri Karray

Despite major advancements in Automatic Speech Recognition (ASR), the state-of-the-art ASR systems struggle to deal with impaired speech even with high-resource languages. In Arabic, this challenge gets amplified, with a…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches

2025-08-11 · Ahmed Aboeitta, Ahmed Sharshar, Youssef Nafea, Shady Shehata arxiv

Speech Recognition (ASR) due to phoneme distortions and high variability. While self-supervised ASR models like Wav2Vec, HuBERT, and Whisper have shown promise, their effectiveness in dysarthric speech remains unclear. T…

Speech Recognition

Adapting Foundation ASR Models to Dysarthric Speech: A Case Study

2026-06-30 · Christian Huber, Laura Kernahan, Alexander Waibel arxiv

Automatic speech recognition (ASR) systems often perform poorly in dysarthric speech, limiting their usefulness to affected speakers in everyday communication. This paper presents a personalized ASR system for a dysarthr…

Speech Recognition

Investigating the Effects of Diffusion-based Conditional Generative Speech Models Used for Speech Enhancement on Dysarthric Speech

2024-12-18 · Joanna Reszka, Parvaneh Janbakhshi, Tilak Purohit, Sadegh Mohammadi

In this study, we aim to explore the effect of pre-trained conditional generative speech models for the first time on dysarthric speech due to Parkinson's disease recorded in an ideal/non-noisy condition. Considering one…

Speech Enhancement

Enhancing AAC Software for Dysarthric Speakers in e-Health Settings: An Evaluation Using TORGO

2024-11-01 · Macarious Hui, Jinda Zhang, Aanchan Mohan

Individuals with cerebral palsy (CP) and amyotrophic lateral sclerosis (ALS) frequently face challenges with articulation, leading to dysarthria and resulting in atypical speech patterns. In healthcare settings, communic…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModellingLarge Language Model+2