Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
Automatic pathological speech detection approaches have shown promising results, gaining attention as potential diagnostic tools alongside costly traditional methods. While these approaches can achieve high accuracy, their lack of interpretability limits their applicability in clinical practice. In this paper, we investigate the use of multimodal Large Language Models (LLMs), specifically ChatGPT-4o, for automatic pathological speech detection in a few-shot in-context learning setting. Experimental results show that this approach not only delivers promising performance but also provides explanations for its decisions, enhancing model interpretability. To further understand its effectiveness, we conduct an ablation study to analyze the impact of different factors, such as input type and system prompts, on the final results. Our findings highlight the potential of multimodal LLMs for further exploration and advancement in automatic pathological speech detection.
Code (0)
등록된 구현이 없습니다.
Tasks
DiagnosticIn-Context LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Exploring the Integration of Large Language Models into Automatic Speech Recognition Systems: An Empirical Study
This paper explores the integration of Large Language Models (LLMs) into Automatic Speech Recognition (ASR) systems to improve transcription accuracy. The increasing sophistication of LLMs, with their in-context learning…
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)In-Context LearningInstruction Following+2Exploring the Capabilities of ChatGPT in Ancient Chinese Translation and Person Name Recognition
ChatGPT's proficiency in handling modern standard languages suggests potential for its use in understanding ancient Chinese. This paper explores ChatGPT's capabilities on ancient Chinese via two tasks: translating ancien…
TranslationDeep Learning for Pathological Speech: A Survey
Advancements in spoken language technologies for neurodegenerative speech disorders are crucial for meeting both clinical and technological needs. This overview paper is vital for advancing the field, as it presents a co…
Automatic Speech RecognitionData AugmentationDeep Learningspeech-recognition+2ChatGPT-EDSS: Empathetic Dialogue Speech Synthesis Trained from ChatGPT-derived Context Word Embeddings
We propose ChatGPT-EDSS, an empathetic dialogue speech synthesis (EDSS) method using ChatGPT for extracting dialogue context. ChatGPT is a chatbot that can deeply understand the content and purpose of an input prompt and…
ChatbotReading ComprehensionSpeech SynthesisWord EmbeddingsIs ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation
ChatGPT, a large-scale language model based on the advanced GPT-3.5 architecture, has shown remarkable potential in various Natural Language Processing (NLP) tasks. However, there is currently a dearth of comprehensive s…
Grammatical Error CorrectionIn-Context LearningLanguage ModelingLanguage Modelling+1