paper-with-me

Papers

Latent Phrase Matching for Dysarthric Speech

2023-06-08 · Colin Lea, Dianna Yee, Jaya Narain, Zifang Huang, Lauren Tooley, Jeffrey P. Bigham, Leah Findlater

Many consumer speech recognition systems are not tuned for people with speech disabilities, resulting in poor recognition and user experience, especially for severe speech differences. Recent studies have emphasized interest in personalized speech models from people with atypical speech patterns. We propose a query-by-example-based personalized phrase recognition system that is trained using small amounts of speech, is language agnostic, does not assume a traditional pronunciation lexicon, and generalizes well across speech difference severities. On an internal dataset collected from 32 people with dysarthria, this approach works regardless of severity and shows a 60% improvement in recall relative to a commercial speech recognition system. On the public EasyCall dataset of dysarthric speech, our approach improves accuracy by 30.5%. Performance degrades as the number of phrases increases, but consistently outperforms ASR systems when trained with 50 unique phrases.

📄 PDF Abstract BibTeX arXiv:2306.05446

Code (0)

등록된 구현이 없습니다.

Tasks

speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Interpretable Deep Learning Model for the Detection and Reconstruction of Dysarthric Speech

2019-07-10 · Daniel Korzekwa, Roberto Barra-Chicote, Bozena Kostek, Thomas Drugman 외

This paper proposed a novel approach for the detection and reconstruction of dysarthric speech. The encoder-decoder model factorizes speech into a low-dimensional latent space and encoding of the input text. We showed th…

Decoder

Variational Auto-Encoder Based Variability Encoding for Dysarthric Speech Recognition

2022-01-24 · Xurong Xie, Rukiye Ruzi, Xunying Liu, Lan Wang

Dysarthric speech recognition is a challenging task due to acoustic variability and limited amount of available data. Diverse conditions of dysarthric speakers account for the acoustic variability, which make the variabi…

speech-recognitionSpeech Recognition

DARS: Dysarthria-Aware Rhythm-Style Synthesis for ASR Enhancement

2026-03-02 · Minghui Wu, Xueling Liu, Jiahuan Fan, Haitao Tang 외 arxiv

Dysarthric speech exhibits abnormal prosody and significant speaker variability, presenting persistent challenges for automatic speech recognition (ASR). While text-to-speech (TTS)-based data augmentation has shown poten…

Speech RecognitionData Augmentation

Prototype-Based Disentanglement for Controllable Dysarthric Speech Synthesis

2026-02-09 · Haoshen Wang, Xueli Zhong, Bingbing Lin, Jia Huang 외 arxiv

Dysarthric speech exhibits high variability and limited labeled data, posing major challenges for both automatic speech recognition (ASR) and assistive speech technologies. Existing approaches rely on synthetic data augm…

Speech RecognitionData AugmentationSpeech Synthesis

Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching

2025-06-19 · Shoutrik Das, Nishant Singh, Arjun Gangwar, S Umesh

Dysarthria is a neurological disorder that significantly impairs speech intelligibility, often rendering affected individuals unable to communicate effectively. This necessitates the development of robust dysarthric-to-r…

Self-Supervised Learning