paper-with-me

홈 › Papers

Voice Disorder Detection Using Long Short Term Memory (LSTM) Model

2018-12-04 · Vibhuti Gupta

Automated detection of voice disorders with computational methods is a recent research area in the medical domain since it requires a rigorous endoscopy for the accurate diagnosis. Efficient screening methods are required for the diagnosis of voice disorders so as to provide timely medical facilities in minimal resources. Detecting Voice disorder using computational methods is a challenging problem since audio data is continuous due to which extracting relevant features and applying machine learning is hard and unreliable. This paper proposes a Long short term memory model (LSTM) to detect pathological voice disorders and evaluates its performance in a real 400 testing samples without any labels. Different feature extraction methods are used to provide the best set of features before applying LSTM model for classification. The paper describes the approach and experiments that show promising results with 22% sensitivity, 97% specificity and 56% unweighted average recall.

📄 PDF Abstract BibTeX arXiv:1812.01779

Code (0)

등록된 구현이 없습니다.

Tasks

Specificity

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Voice Disorder Analysis: a Transformer-based Approach

2024-06-20 · Alkis Koudounas, Gabriele Ciravegna, Marco Fantini, Giovanni Succo 외

Voice disorders are pathologies significantly affecting patient quality of life. However, non-invasive automated diagnosis of these pathologies is still under-explored, due to both a shortage of pathological voice data, …

Data AugmentationDiversitySentenceSynthetic Data Generation

Continuous Speech for Improved Learning Pathological Voice Disorders

2022-02-22 · Syu-Siang Wang, Chi-Te Wang, Chih-Chung Lai, Yu Tsao 외

Goal: Numerous studies had successfully differentiated normal and abnormal voice samples. Nevertheless, further classification had rarely been attempted. This study proposes a novel approach, using continuous Mandarin sp…

AI-Driven Acoustic Voice Biomarker-Based Hierarchical Classification of Benign Laryngeal Voice Disorders from Sustained Vowels

2025-12-31 · Mohsen Annabestani, Samira Aghadoost, Anais Rameau, Olivier Elemento 외 arxiv

Benign laryngeal voice disorders affect nearly one in five individuals and often manifest as dysphonia, while also serving as non-invasive indicators of broader physiological dysfunction. We introduce a clinically inspir…

Integration of Text and Graph-based Features for Detecting Mental Health Disorders from Voice

2022-05-14 · Nasser Ghadiri, Rasoul Samani, Fahime Shahrokh

With the availability of voice-enabled devices such as smart phones, mental health disorders could be detected and treated earlier, particularly post-pandemic. The current methods involve extracting features directly fro…

Depression Detection

Voice Pathology Detection Using Phonation

2025-08-11 · Sri Raksha Siva, Nived Suthahar, Prakash Boominathan, Uma Ranjan arxiv

Voice disorders significantly affect communication and quality of life, requiring an early and accurate diagnosis. Traditional methods like laryngoscopy are invasive, subjective, and often inaccessible. This research pro…

Voice pathology detectionData Augmentation