paper-with-me

홈 › Papers

Medical Speech Symptoms Classification via Disentangled Representation

2024-03-08 · Jianzong Wang, Pengcheng Li, xulong Zhang, Ning Cheng, Jing Xiao

Intent is defined for understanding spoken language in existing works. Both textual features and acoustic features involved in medical speech contain intent, which is important for symptomatic diagnosis. In this paper, we propose a medical speech classification model named DRSC that automatically learns to disentangle intent and content representations from textual-acoustic data for classification. The intent representations of the text domain and the Mel-spectrogram domain are extracted via intent encoders, and then the reconstructed text feature and the Mel-spectrogram feature are obtained through two exchanges. After combining the intent from two domains into a joint representation, the integrated intent representation is fed into a decision layer for classification. Experimental results show that our model obtains an average accuracy rate of 95% in detecting 25 different medical symptoms.

📄 PDF Abstract BibTeX arXiv:2403.05000

Code (0)

등록된 구현이 없습니다.

Tasks

Classification

Similar Papers 제목 키워드 기반

Equine Pain Behavior Classification via Self-Supervised Disentangled Pose Representation

2021-08-30 · Maheen Rashid, Sofia Broomé, Katrina Ask, Elin Hernlund 외

Timely detection of horse pain is important for equine welfare. Horses express pain through their facial and body behavior, but may hide signs of pain from unfamiliar human observers. In addition, collecting visual data …

Classification

MultiQT: Multimodal Learning for Real-Time Question Tracking in Speech

2020-05-02 · ACL 2020 6 · Jakob D. Havtorn, Jan Latko, Joakim Edin, Lasse Borgholt 외

We address a challenging and practical task of labeling questions in speech in real time during telephone calls to emergency medical services in English, which embeds within a broader decision support system for emergenc…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

An Ensemble Classification Approach in A Multi-Layered Large Language Model Framework for Disease Prediction

2025-09-02 · Ali Hamdi, Malak Mohamed, Rokaia Emad, Khaled Shaban arxiv

Social telehealth has made remarkable progress in healthcare by allowing patients to post symptoms and participate in medical consultations remotely. Users frequently post symptoms on social media and online health platf…

Ensemble Learning

Towards Learning Fine-Grained Disentangled Representations from Speech

2018-08-08 · Yuan Gong, Christian Poellabauer

Learning disentangled representations of high-dimensional data is currently an active research area. However, compared to the field of computer vision, less work has been done for speech processing. In this paper, we pro…

Representation LearningSpeech Representation Learning

Protecting gender and identity with disentangled speech representations

2021-04-22 · Dimitrios Stoidis, Andrea Cavallaro

Besides its linguistic content, our speech is rich in biometric information that can be inferred by classifiers. Learning privacy-preserving representations for speech signals enables downstream tasks without sharing unn…

Privacy PreservingRepresentation LearningSpeaker VerificationSpeech Recognition