paper-with-me

Papers

Constrained speaker linking

2014-03-26 · David A. van Leeuwen, Niko Brümmer

In this paper we study speaker linking (a.k.a.\ partitioning) given constraints of the distribution of speaker identities over speech recordings. Specifically, we show that the intractable partitioning problem becomes tractable when the constraints pre-partition the data in smaller cliques with non-overlapping speakers. The surprisingly common case where speakers in telephone conversations are known, but the assignment of channels to identities is unspecified, is treated in a Bayesian way. We show that for the Dutch CGN database, where this channel assignment task is at hand, a lightweight speaker recognition system can quite effectively solve the channel assignment problem, with 93% of the cliques solved. We further show that the posterior distribution over channel assignment configurations is well calibrated.

📄 PDF Abstract BibTeX arXiv:1403.7084

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Recognition

Similar Papers 제목 키워드 기반

Large-Scale Speaker Diarization of Radio Broadcast Archives

2019-06-19 · Emre Yilmaz, Adem Derinel, Zhou Kun, Henk van den Heuvel 외

This paper describes our initial efforts to build a large-scale speaker diarization (SD) and identification system on a recently digitized radio broadcast archive from the Netherlands which has more than 6500 audio tapes…

speaker-diarizationSpeaker DiarizationSpeaker Identification

Schwa-deletion in German noun-noun compounds

2020-12-01 · COLING (CogALex) 2020 12 · Tom S Juzek, Jana Haeussler

We report ongoing research on linking elements in German compounds, with a focus on noun-noun compounds in which the first constituent is ending in schwa. We present a corpus of about 3000 nouns ending in schwa, annotate…

Evaluating the Effects of Embedding with Speaker Identity Information in Dialogue Summarization

2022-06-01 · LREC 2022 6 · Yuji Naraki, Tetsuya Sakai, Yoshihiko Hayashi

Automatic dialogue summarization is a task used to succinctly summarize a dialogue transcript while correctly linking the speakers and their speech, which distinguishes this task from a conventional document summarizatio…

Document SummarizationInformativenessPosition

G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition

2026-03-11 · Jing Peng, Ziyi Chen, Haoyu Li, Yucheng Wang 외 arxiv

We study timestamped speaker-attributed automatic speech recognition (SA-ASR) for long-form, multi-party speech with overlap. In this setting, chunk-wise inference must preserve meeting-level speaker identity consistency…

Speech Recognition

DG^VoiC: Speaker Clustering for Fraud Investigation under Real Call-Centre Conditions

2026-06-26 · Muhammad Shakeel Akram, Amal Htait, Abdul Hamid Sadka, Emma Meisingseth 외 arxiv

Insurance fraud remains costly and operationally difficult, particularly in call-centre workflows where many customer interactions begin at FNOL. While recent fraud detection methods mainly rely on structured data, text,…

Fraud Detection