paper-with-me

홈 › Papers

Learning to detect dysarthria from raw speech

2018-11-27 · Juliette Millet, Neil Zeghidour

Speech classifiers of paralinguistic traits traditionally learn from diverse hand-crafted low-level features, by selecting the relevant information for the task at hand. We explore an alternative to this selection, by learning jointly the classifier, and the feature extraction. Recent work on speech recognition has shown improved performance over speech features by learning from the waveform. We extend this approach to paralinguistic classification and propose a neural network that can learn a filterbank, a normalization factor and a compression power from the raw speech, jointly with the rest of the architecture. We apply this model to dysarthria detection from sentence-level audio recordings. Starting from a strong attention-based baseline on which mel-filterbanks outperform standard low-level descriptors, we show that learning the filters or the normalization and compression improves over fixed features by 10% absolute accuracy. We also observe a gain over OpenSmile features by learning jointly the feature extraction, the normalization, and the compression factor with the architecture. This constitutes a first attempt at learning jointly all these operations from raw audio for a speech classification task.

📄 PDF Abstract BibTeX arXiv:1811.11101

Code (3)

FastAndFourier/MLA-Project-DYSARTHRIA tf
dkman94/FastrApp-Android
fastrapp/fastr-android-app

Tasks

General ClassificationSentencespeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Interpretable Deep Learning Model for the Detection and Reconstruction of Dysarthric Speech

2019-07-10 · Daniel Korzekwa, Roberto Barra-Chicote, Bozena Kostek, Thomas Drugman 외

This paper proposed a novel approach for the detection and reconstruction of dysarthric speech. The encoder-decoder model factorizes speech into a low-dimensional latent space and encoding of the input text. We showed th…

Decoder

Brain Signals to Rescue Aphasia, Apraxia and Dysarthria Speech Recognition

2021-02-28 · Gautam Krishna, Mason Carnahan, Shilpa Shamapant, Yashitha Surendranath 외

In this paper, we propose a deep learning-based algorithm to improve the performance of automatic speech recognition (ASR) systems for aphasia, apraxia, and dysarthria speech by utilizing electroencephalography (EEG) fea…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)EEGElectroencephalogram (EEG)+2

Wav2vec-based Detection and Severity Level Classification of Dysarthria from Speech

2023-09-25 · Farhad Javanmardi, Saska Tirronen, Manila Kodali, Sudarsana Reddy Kadiri 외

Automatic detection and severity level classification of dysarthria directly from acoustic speech signals can be used as a tool in medical diagnosis. In this work, the pre-trained wav2vec 2.0 model is studied as a featur…

ClassificationMedical Diagnosis

Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease

2026-03-23 · Abner Hernandez, Eunjung Yeo, Kwanghee Choi, Chin-Jou Li 외 arxiv

The limited availability of dysarthric speech data makes cross-lingual detection an important but challenging problem. A key difficulty is that speech representations often encode language-dependent structure that can co…

Unsupervised Domain Adaptation for Dysarthric Speech Detection via Domain Adversarial Training and Mutual Information Minimization

2021-06-18 · Disong Wang, Liqun Deng, Yu Ting Yeung, Xiao Chen 외

Dysarthric speech detection (DSD) systems aim to detect characteristics of the neuromotor disorder from speech. Such systems are particularly susceptible to domain mismatch where the training and testing data come from t…

Domain AdaptationMulti-Task LearningUnsupervised Domain Adaptation