paper-with-me

Papers

Pathological speech detection using x-vector embeddings

2020-03-02 · Catarina Botelho, Francisco Teixeira, Thomas Rolland, Alberto Abad, Isabel Trancoso

The potential of speech as a non-invasive biomarker to assess a speaker's health has been repeatedly supported by the results of multiple works, for both physical and psychological conditions. Traditional systems for speech-based disease classification have focused on carefully designed knowledge-based features. However, these features may not represent the disease's full symptomatology, and may even overlook its more subtle manifestations. This has prompted researchers to move in the direction of general speaker representations that inherently model symptoms, such as Gaussian Supervectors, i-vectors and, x-vectors. In this work, we focus on the latter, to assess their applicability as a general feature extraction method to the detection of Parkinson's disease (PD) and obstructive sleep apnea (OSA). We test our approach against knowledge-based features and i-vectors, and report results for two European Portuguese corpora, for OSA and PD, as well as for an additional Spanish corpus for PD. Both x-vector and i-vector models were trained with an out-of-domain European Portuguese corpus. Our results show that x-vectors are able to perform better than knowledge-based features in same-language corpora. Moreover, while x-vectors performed similarly to i-vectors in matched conditions, they significantly outperform them when domain-mismatch occurs.

📄 PDF Abstract BibTeX arXiv:2003.00864

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection

2024-09-13 · Yacouba Kaloga, Shakeel A. Sheikh, Ina Kodrasi

Recently proposed automatic pathological speech detection approaches rely on spectrogram input representations or wav2vec2 embeddings. These representations may contain pathology irrelevant uncorrelated information, such…

Dimensionality Reduction

Selfsupervised learning for pathological speech detection

2024-05-16 · Shakeel Ahmad Sheikh

Speech production is a complex phenomenon, wherein the brain orchestrates a sequence of processes involving thought processing, motor planning, and the execution of articulatory movements. However, this intricate executi…

Self-Supervised Learning

Impact of Speech Mode in Automatic Pathological Speech Detection

2024-06-14 · Shakeel A. Sheikh, Ina Kodrasi

Automatic pathological speech detection approaches yield promising results in identifying various pathologies. These approaches are typically designed and evaluated for phonetically-controlled speech scenarios, where spe…

Navigate

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

2026-06-04 · Fernando López, Santosh Kesiraju, Jordi Luque arxiv

Automatic speech recognition (ASR) has advanced remarkably for standard speech; however, pathological speech from neurological conditions remains a significant challenge. We investigate speaker conditioning via Feature-w…

parameter-efficient fine-tuningSpeech Recognition

Glottal Closure Instants Detection From Pathological Acoustic Speech Signal Using Deep Learning

2018-11-25 · Gurunath Reddy M, Tanumay Mandal, Krothapalli Sreenivasa Rao

In this paper, we propose a classification based glottal closure instants (GCI) detection from pathological acoustic speech signal, which finds many applications in vocal disorder analysis. Till date, GCI for pathologica…

General Classification