Multilingual Stutter Event Detection for English, German, and Mandarin Speech
This paper presents a multi-label stuttering detection system trained on multi-corpus, multilingual data in English, German, and Mandarin.By leveraging annotated stuttering data from three languages and four corpora, the model captures language-independent characteristics of stuttering, enabling robust detection across linguistic contexts. Experimental results demonstrate that multilingual training achieves performance comparable to and, in some cases, even exceeds that of previous systems. These findings suggest that stuttering exhibits cross-linguistic consistency, which supports the development of language-agnostic detection systems. Our work demonstrates the feasibility and advantages of using multilingual data to improve generalizability and reliability in automated stuttering detection.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
On the Difficulty of Token-Level Modeling of Dysfluency and Fluency Shaping Artifacts
Automatic transcription of stuttered speech remains a challenge, even for modern end-to-end (E2E) automatic speech recognition (ASR) frameworks. Dysfluencies and fluency-shaping artifacts are often overlooked, resulting …
Speech RecognitionA Stutter Seldom Comes Alone -- Cross-Corpus Stuttering Detection as a Multi-label Problem
Most stuttering detection and classification research has viewed stuttering as a multi-class classification problem or a binary detection task for each dysfluency type; however, this does not match the nature of stutteri…
ClassificationCross-corpusMulti-class ClassificationMulti-Task LearningDysfluencies Seldom Come Alone -- Detection as a Multi-Label Problem
Specially adapted speech recognition models are necessary to handle stuttered speech. For these to be used in a targeted manner, stuttered speech must be reliably detected. Recent works have treated stuttering as a multi…
Multi-class Classificationspeech-recognitionSpeech RecognitionDetecting Dysfluencies in Stuttering Therapy Using wav2vec 2.0
Stuttering is a varied speech disorder that harms an individual's communication ability. Persons who stutter (PWS) often use speech therapy to cope with their condition. Improving speech recognition systems for people wi…
Multi-Task Learningspeech-recognitionSpeech RecognitionLarge Language Models for Dysfluency Detection in Stuttered Speech
Accurately detecting dysfluencies in spoken language can help to improve the performance of automatic speech and language processing components and support the development of more inclusive speech and language technologi…
Automatic Speech RecognitionLanguage ModelingLanguage Modellingspeech-recognition+1