paper-with-me

홈 › Papers

Something from Nothing: Data Augmentation for Robust Severity Level Estimation of Dysarthric Speech

2026-03-16 · Jaesung Bae, Xiuwen Zheng, Minje Kim, Chang D. Yoo, Mark Hasegawa-Johnson arxiv

Dysarthric speech quality assessment (DSQA) is critical for clinical diagnostics and inclusive speech technologies. However, subjective evaluation is costly and difficult to scale, and the scarcity of labeled data limits robust objective modeling. To address this, we propose a three-stage framework that leverages unlabeled dysarthric speech and large-scale typical speech datasets to scale training. A teacher model first generates pseudo-labels for unlabeled samples, followed by weakly supervised pretraining using a label-aware contrastive learning strategy that exposes the model to diverse speakers and acoustic conditions. The pretrained model is then fine-tuned for the downstream DSQA task. Experiments on five unseen datasets spanning multiple etiologies and languages demonstrate the robustness of our approach. Our Whisper-based baseline significantly outperforms SOTA DSQA predictors such as SpICE, and the full framework achieves an average SRCC of 0.761 across unseen test datasets.

📄 PDF Abstract BibTeX arXiv:2603.15988

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningData Augmentation

Similar Papers 제목 키워드 기반

Disease Severity Regression with Continuous Data Augmentation

2023-02-24 · Shumpei Takezaki, Kiyohito Tanaka, Seiichi Uchida, Takeaki Kadota

Disease severity regression by a convolutional neural network (CNN) for medical images requires a sufficient number of image samples labeled with severity levels. Conditional generative adversarial network (cGAN)-based d…

Data AugmentationGenerative Adversarial Networkregression

Improving End-to-End Speech Recognition for Dysarthric Speech through In-Domain Data Augmentation

2026-06-18 · Paban Sapkota, Hemant Kumar Kathania, Sudarsana Reddy Kadiri, Shrikanth Narayanan arxiv

Dysarthric speech recognition is crucial for facilitating effective communication among individuals with dysarthria. However, accurately recognizing dysarthric speech poses significant challenges due to varying severity …

Speech RecognitionData Augmentation

Enhancing Knee Osteoarthritis severity level classification using diffusion augmented images

2023-09-17 · Paleti Nikhil Chowdary, Gorantla V N S L Vishnu Vardhan, Menta Sai Akshay, Menta Sai Aashish 외

This research paper explores the classification of knee osteoarthritis (OA) severity levels using advanced computer vision models and augmentation techniques. The study investigates the effectiveness of data preprocessin…

Data Augmentation

Retro-Actions: Learning 'Close' by Time-Reversing 'Open' Videos

2019-09-20 · Will Price, Dima Damen

We investigate video transforms that result in class-homogeneous label-transforms. These are video transforms that consistently maintain or modify the labels of all videos in each class. We propose a general approach to …

Data AugmentationVideo RecognitionZero-Shot Learning

Accurate synthesis of Dysarthric Speech for ASR data augmentation

2023-08-16 · Mohammad Soleymanpour, Michael T. Johnson, Rahim Soleymanpour, Jeffrey Berry

Dysarthria is a motor speech disorder often characterized by reduced speech intelligibility through slow, uncoordinated control of speech production muscles. Automatic Speech recognition (ASR) systems can help dysarthric…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+2