paper-with-me

홈 › Papers

Domain-Aware Speaker Diarization On African-Accented English

2025-09-25 · Chibuzor Okocha, Kelechi Ezema, Christan Grant arxiv

This study examines domain effects in speaker diarization for African-accented English. We evaluate multiple production and open systems on general and clinical dialogues under a strict DER protocol that scores overlap. A consistent domain penalty appears for clinical speech and remains significant across models. Error analysis attributes much of this penalty to false alarms and missed detections, aligning with short turns and frequent overlap. We test lightweight domain adaptation by fine-tuning a segmentation module on accent-matched data; it reduces error but does not eliminate the gap. Our contributions include a controlled benchmark across domains, a concise approach to error decomposition and conversation-level profiling, and an adaptation recipe that is easy to reproduce. Results point to overlap-aware segmentation and balanced clinical resources as practical next steps.

📄 PDF Abstract BibTeX arXiv:2509.21554

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker DiarizationDomain Adaptation

Similar Papers 제목 키워드 기반

Afrispeech-Dialog: A Benchmark Dataset for Spontaneous English Conversations in Healthcare and Beyond

2025-02-06 · Mardhiyah Sanni, Tassallah Abdullahi, Devendra D. Kayande, Emmanuel Ayodele 외

Speech technologies are transforming interactions across various sectors, from healthcare to call centers and robots, yet their performance on African-accented conversations remains underexplored. We introduce Afrispeech…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Conversation Summarizationspeaker-diarization+3

1000 African Voices: Advancing inclusive multi-speaker multi-accent speech synthesis

2024-06-17 · Sewade Ogun, Abraham T. Owodunni, Tobi Olatunji, Eniola Alese 외

Recent advances in speech synthesis have enabled many useful applications like audio directions in Google Maps, screen readers, and automated content generation on platforms like TikTok. However, these systems are mostly…

DiversitySpeech Synthesis

AfriSpeech-200: Pan-African Accented Speech Dataset for Clinical and General Domain ASR

2023-09-30 · Tobi Olatunji, Tejumade Afonja, Aditya Yadavalli, Chris Chinenye Emezue 외

Africa has a very low doctor-to-patient ratio. At very busy clinics, doctors could see 30+ patients per day -- a heavy patient burden compared with developed countries -- but productivity tools such as clinical automatic…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

TOLD: A Novel Two-Stage Overlap-Aware Framework for Speaker Diarization

2023-03-08 · JiaMing Wang, Zhihao Du, Shiliang Zhang

Recently, end-to-end neural diarization (EEND) is introduced and achieves promising results in speaker-overlapped scenarios. In EEND, speaker diarization is formulated as a multi-label prediction problem, where speaker a…

speaker-diarizationSpeaker DiarizationVocal Bursts Valence Prediction

Community Detection Graph Convolutional Network for Overlap-Aware Speaker Diarization

2023-06-26 · Jie Wang, Zhicong Chen, Haodong Zhou, Lin Li 외

The clustering algorithm plays a crucial role in speaker diarization systems. However, traditional clustering algorithms suffer from the complex distribution of speaker embeddings and lack of digging potential relationsh…

ClusteringCommunity DetectionGraph Generationspeaker-diarization+1