Speaker Diarization
12개 벤치마크 · 논문 379편 · 이 태스크의 논문 보기 →
Benchmarks
CALLHOME
NIST-SRE 2000
AMI Lapel
AMI MixHeadset
CH109
DIHARD
ETAPE
AMI
CALLHOME-109
AliMeeting
DIHARD II
Hub5'00 CallHome
Most implemented
AVA-AVD: Audio-Visual Speaker Diarization in the Wild
VibeVoice-ASR-Streaming Technical Report
Speaker Diarization with LSTM
Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech
Speech Emotion Diarization: Which Emotion Appears When?
The Third DIHARD Diarization Challenge
Papers
VibeVoice-ASR-Streaming Technical Report
Traditional speaker-attributed ASR systems treated ASR and speaker diarization as two separate tasks. Recently, end-to-end models such as VibeVoice-ASR have unified the two tasks within a single model. However, existing …
Speaker DiarizationTowards Actionable Surgical Team Dynamics: from Teamwork to Counterfactual Annotations
Modeling team interactions in high-stakes environments such as operating rooms is critical for understanding how coordination, communication, and individual behaviors shape team performance and safety outcomes. Existing …
Speaker DiarizationLeading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM Challenge
The second Multilingual Conversational Speech Language Model (MLC-SLM) Challenge evaluates two tasks over complete, unsegmented multilingual conversations: speaker diarization and recognition (Task 1) and conversational …
Speaker DiarizationQuantifying the Sources of Instability in LLM-Based Stance Analysis of Public Discourse
Computational social science increasingly relies on automated preprocessing pipelines -- speaker diarization, ASR transcript cleaning, sentence segmentation -- to convert raw media into analyzable text. When these pipeli…
Speaker DiarizationDiarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech
This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system combines a modular speaker diarization front end with a challenge-adap…
Reinforcement LearningSpeaker DiarizationActivity DetectionAutomatic Detection of Stress from Speech in the Trier Social Stress Test
Automatically detecting stress in speech provides an unobtrusive way to gain insights relevant to behavioral research or clinical assessment. This study investigates the automatic differentiation between a stressful and …
Speaker Diarization