paper-with-me

홈 › Papers

Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition

2025-09-23 · Niclas Pokel, Pehuén Moure, Roman Boehringer, Shih-Chii Liu, Yingqiang Gao arxiv

Speech impairments resulting from congenital disorders, such as cerebral palsy, down syndrome, or apert syndrome, as well as acquired brain injuries due to stroke, traumatic accidents, or tumors, present major challenges to automatic speech recognition (ASR) systems. Despite recent advancements, state-of-the-art ASR models like Whisper still struggle with non-normative speech due to limited training data availability and high acoustic variability. Moreover, collecting and annotating non-normative speech is burdensome: speaking is effortful for many affected individuals, while laborious annotation often requires caregivers familiar with the speaker. This work introduces a novel ASR personalization method based on Bayesian Low-rank Adaptation for data-efficient fine-tuning. We validate our method on the English UA-Speech dataset and a newly collected German speech dataset, BF-Sprache, from a child with structural speech impairment. The dataset and approach are designed to reflect the challenges of low-resource settings that include individuals with speech impairments. Our method significantly improves ASR accuracy for impaired speech while maintaining data and annotation efficiency, offering a practical path toward inclusive ASR.

📄 PDF Abstract BibTeX arXiv:2509.20397

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Recognition

Similar Papers 제목 키워드 기반

Adversarial Data Augmentation Using VAE-GAN for Disordered Speech Recognition

2022-11-03 · Zengrui Jin, Xurong Xie, Mengzhe Geng, Tianzi Wang 외

Automatic recognition of disordered speech remains a highly challenging task to date. The underlying neuro-motor conditions, often compounded with co-occurring physical disabilities, lead to the difficulty in collecting …

Data AugmentationGenerative Adversarial Networkspeech-recognitionSpeech Recognition

VoiceBank-2023: A Multi-Speaker Mandarin Speech Corpus for Constructing Personalized TTS Systems for the Speech Impaired

2023-08-27 · Jia-Jyu Su, Pang-Chen Liao, Yen-Ting Lin, Wu-Hao Li 외

Services of personalized TTS systems for the Mandarin-speaking speech impaired are rarely mentioned. Taiwan started the VoiceBanking project in 2020, aiming to build a complete set of services to deliver personalized Man…

Data-Efficient ASR Personalization for Non-Normative Speech Using an Uncertainty-Based Phoneme Difficulty Score for Guided Sampling

2025-09-23 · Niclas Pokel, Pehuén Moure, Roman Böhringer, Yingqiang Gao arxiv

ASR systems struggle with non-normative speech due to high acoustic variability and data scarcity. We propose a data-efficient method using phoneme-level uncertainty to guide fine-tuning for personalization. Instead of c…

An analysis of degenerating speech due to progressive dysarthria on ASR performance

2022-10-31 · Katrin Tomanek, Katie Seaver, Pan-Pan Jiang, Richard Cave 외

Although personalized automatic speech recognition (ASR) models have recently been designed to recognize even severely impaired speech, model performance may degrade over time for persons with degenerating speech. The ai…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Gated Low-rank Adaptation for personalized Code-Switching Automatic Speech Recognition on the low-spec devices

2024-04-24 · Gwantae Kim, Bokyeung Lee, Donghyeon Kim, Hanseok Ko

In recent times, there has been a growing interest in utilizing personalized large models on low-spec devices, such as mobile and CPU-only devices. However, utilizing a personalized large model in the on-device is ineffi…

Automatic Speech RecognitionCPUparameter-efficient fine-tuningspeech-recognition+1