paper-with-me

홈 › Papers

Combining Prosodic, Voice Quality and Lexical Features to Automatically Detect Alzheimer's Disease

2020-11-18 · Mireia Farrús, Joan Codina-Filbà

Alzheimer's Disease (AD) is nowadays the most common form of dementia, and its automatic detection can help to identify symptoms at early stages, so that preventive actions can be carried out. Moreover, non-intrusive techniques based on spoken data are crucial for the development of AD automatic detection systems. In this light, this paper is presented as a contribution to the ADReSS Challenge, aiming at improving AD automatic detection from spontaneous speech. To this end, recordings from 108 participants, which are age-, gender-, and AD condition-balanced, have been used as training set to perform two different tasks: classification into AD/non-AD conditions, and regression over the Mini-Mental State Examination (MMSE) scores. Both tasks have been performed extracting 28 features from speech -- based on prosody and voice quality -- and 51 features from the transcriptions -- based on lexical and turn-taking information. Our results achieved up to 87.5 % of classification accuracy using a Random Forest classifier, and 4.54 of RMSE using a linear regression with stochastic gradient descent over the provided test set. This shows promising results in the automatic detection of Alzheimer's Disease through speech and lexical features.

📄 PDF Abstract BibTeX arXiv:2011.09272

Code (0)

등록된 구현이 없습니다.

Tasks

regression

Methods 이 논문이 사용한 방법론

Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Voice Quality and Pitch Features in Transformer-Based Speech Recognition

2021-12-21 · Guillermo Cámbara, Jordi Luque, Mireia Farrús

Jitter and shimmer measurements have shown to be carriers of voice quality and prosodic information which enhance the performance of tasks like speaker recognition, diarization or automatic speech recognition (ASR). Howe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speaker Recognitionspeech-recognition+1

The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations

2026-01-20 · Sam OConnor Russell, Delphine Charuau, Naomi Harte arxiv

Fluid turn-taking remains a key challenge in human-robot interaction. Self-supervised speech representations (S3Rs) have driven many advances, but it remains unclear whether S3R-based turn-taking models rely on prosodic …

What Does a Pathological Speech Assessment Model Know about Acoustic Features? A Case Study on Oral and Oropharyngeal Cancer Patients

2026-06-23 · Tuan Nguyen, Corinne Fredouille, Alain Ghio, Muriel Lalain 외 arxiv

This work investigates the interpretability of a Wav2Vec 2.0based speech intelligibility assessment model for oral and oropharyngeal cancer patients through canonical correlation analysis. By measuring the correlation be…

Respiratory Distress Detection from Telephone Speech using Acoustic and Prosodic Features

2020-11-15 · Meemnur Rashid, Kaisar Ahmed Alman, Khaled Hasan, John H. L. Hansen 외

With the widespread use of telemedicine services, automatic assessment of health conditions via telephone speech can significantly impact public health. This work summarizes our preliminary findings on automatic detectio…

Disentangling segmental and prosodic factors to non-native speech comprehensibility

2024-08-20 · Waris Quamer, Ricardo Gutierrez-Osuna

Current accent conversion (AC) systems do not disentangle the two main sources of non-native accent: segmental and prosodic characteristics. Being able to manipulate a non-native speaker's segmental and/or prosodic chann…

QuantizationVoice Similarity