paper-with-me

Speaker Diarization

12개 벤치마크 · 논문 379편 · 이 태스크의 논문 보기 →

Benchmarks

CALLHOME

결과 11개

NIST-SRE 2000

결과 5개

AMI Lapel

결과 4개

AMI MixHeadset

결과 4개

CH109

결과 4개

DIHARD

결과 3개

ETAPE

결과 3개

AMI

결과 2개

CALLHOME-109

결과 2개

AliMeeting

결과 1개

DIHARD II

결과 1개

Hub5'00 CallHome

결과 1개

Most implemented

VibeVoice-ASR-Streaming Technical Report

2026-09-02 · 구현 7개

Speaker Diarization with LSTM

2017-10-28 · 구현 4개

The Third DIHARD Diarization Challenge

2020-12-02 · 구현 3개

Papers

VibeVoice-ASR-Streaming Technical Report

2026-09-02 · Yujie Tu, Zhiliang Peng, Jianwei Yu, Li Dong 외 hf

Traditional speaker-attributed ASR systems treated ASR and speaker diarization as two separate tasks. Recently, end-to-end models such as VibeVoice-ASR have unified the two tasks within a single model. However, existing …

Speaker Diarization

Towards Actionable Surgical Team Dynamics: from Teamwork to Counterfactual Annotations

2026-08-24 · Vincenzo Marco De Luca, Antonio Longa, Andrea Passerini arxiv

Modeling team interactions in high-stakes environments such as operating rooms is critical for understanding how coordination, communication, and individual behaviors shape team performance and safety outcomes. Existing …

Speaker Diarization

Leading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM Challenge

2026-08-14 · Kexin Shi, Renhe Sun, Yuge Huang, Ximeng Wang 외 arxiv

The second Multilingual Conversational Speech Language Model (MLC-SLM) Challenge evaluates two tasks over complete, unsegmented multilingual conversations: speaker diarization and recognition (Task 1) and conversational …

Speaker Diarization

Quantifying the Sources of Instability in LLM-Based Stance Analysis of Public Discourse

2026-07-12 · Bo Chen arxiv

Computational social science increasingly relies on automated preprocessing pipelines -- speaker diarization, ASR transcript cleaning, sentence segmentation -- to convert raw media into analyzable text. When these pipeli…

Speaker Diarization

Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech

2026-07-09 · Hao Wu, RongQi Han, Zhen Wang, Wei Liang 외 arxiv

This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system combines a modular speaker diarization front end with a challenge-adap…

Reinforcement LearningSpeaker DiarizationActivity Detection

Automatic Detection of Stress from Speech in the Trier Social Stress Test

2026-07-01 · Hanna Drimalla, Wieland R. Cremer, Christine Kraus, Oliver T. Wolf arxiv

Automatically detecting stress in speech provides an unobtrusive way to gain insights relevant to behavioral research or clinical assessment. This study investigates the automatic differentiation between a stressful and …

Speaker Diarization

전체 379편 보기 →