paper-with-me

홈 › Papers

Confidence-Guided Error Correction for Disordered Speech Recognition

2025-09-29 · Abner Hernandez, Tomás Arias Vergara, Andreas Maier, Paula Andrea Pérez-Toro arxiv

We investigate the use of large language models (LLMs) as post-processing modules for automatic speech recognition (ASR), focusing on their ability to perform error correction for disordered speech. In particular, we propose confidence-informed prompting, where word-level uncertainty estimates are embedded directly into LLM training to improve robustness and generalization across speakers and datasets. This approach directs the model to uncertain ASR regions and reduces overcorrection. We fine-tune a LLaMA 3.1 model and compare our approach to both transcript-only fine-tuning and post hoc confidence-based filtering. Evaluations show that our method achieves a 10% relative WER reduction compared to naive LLM correction on the Speech Accessibility Project spontaneous speech and a 47% reduction on TORGO, demonstrating the effectiveness of confidence-aware fine-tuning for impaired speech.

📄 PDF Abstract BibTeX arXiv:2509.25048

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

Learnings from curating a trustworthy, well-annotated, and useful dataset of disordered English speech

2024-09-13 · Pan-Pan Jiang, Jimmy Tobin, Katrin Tomanek, Robert L. MacDonald 외

Project Euphonia, a Google initiative, is dedicated to improving automatic speech recognition (ASR) of disordered speech. A central objective of the project is to create a large, high-quality, and diverse speech corpus. …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Error Correction by Paying Attention to Both Acoustic and Confidence References for Automatic Speech Recognition

2024-06-29 · Yuchun Shu, Bo Hu, Yifeng He, Hao Shi 외

Accurately finding the wrong words in the automatic speech recognition (ASR) hypothesis and recovering them well-founded is the goal of speech error correction. In this paper, we propose a non-autoregressive speech error…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Evaluating ASR Confidence Scores for Automated Error Detection in User-Assisted Correction Interfaces

2025-03-19 · Korbinian Kuhn, Verena Kersken, Gottfried Zimmermann

Despite advances in Automatic Speech Recognition (ASR), transcription errors persist and require manual correction. Confidence scores, which indicate the certainty of ASR results, could assist users in identifying and co…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Adversarial Data Augmentation for Disordered Speech Recognition

2021-08-02 · Zengrui Jin, Mengzhe Geng, Xurong Xie, Jianwei Yu 외

Automatic recognition of disordered speech remains a highly challenging task to date. The underlying neuro-motor conditions, often compounded with co-occurring physical disabilities, lead to the difficulty in collecting …

Data Augmentationspeech-recognitionSpeech Recognition

Investigation of Data Augmentation Techniques for Disordered Speech Recognition

2022-01-14 · Mengzhe Geng, Xurong Xie, Shansong Liu, Jianwei Yu 외

Disordered speech recognition is a highly challenging task. The underlying neuro-motor conditions of people with speech disorders, often compounded with co-occurring physical disabilities, lead to the difficulty in colle…

Data Augmentationspeech-recognitionSpeech Recognition