paper-with-me

홈 › Papers

Cross-Lingual Speaker Identification from Weak Local Evidence

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Speaker identification, determining which character said each utterance in text, benefits many downstream tasks. Most existing approaches use expert-defined rules or rule-based features to directly approach this task, but these approaches come with significant drawbacks, such as lack of contextual reasoning and poor cross-lingual generalization. In this work, we propose a speaker identification framework that addresses these issues. We first extract large-scale distant supervision signals in English via general-purpose tools and heuristics, and then apply these weakly-labeled instances with a focus on encouraging contextual reasoning to train a cross-lingual language model. We show that our final model outperforms the previous state-of-the-art methods on two English speaker identification benchmarks by $5.4\%$ in accuracy, as well as two Chinese speaker identification datasets by up to $4.7\%$.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingSpeaker Identification

Similar Papers 제목 키워드 기반

Cross-Lingual Speaker Identification Using Distant Supervision

2022-10-11 · Ben Zhou, Dian Yu, Dong Yu, Dan Roth

Speaker identification, determining which character said each utterance in literary text, benefits many downstream tasks. Most existing approaches use expert-defined rules or rule-based features to directly approach this…

Language ModelingLanguage ModellingSpeaker Identification

Whisper Speaker Identification: Leveraging Pre-Trained Multilingual Transformers for Robust Speaker Embeddings

2025-03-13 · Jakaria Islam Emon, Md Abu Salek, Kazi Tamanna Alam

Speaker identification in multilingual settings presents unique challenges, particularly when conventional models are predominantly trained on English data. In this paper, we propose WSI (Whisper Speaker Identification),…

Speaker Identificationspeech-recognitionSpeech Recognition

POLY-SIM: Polyglot Speaker Identification with Missing Modality Grand Challenge 2026 Evaluation Plan

2026-03-25 · Marta Moscati, Muhammad Saad Saeed, Marina Zanoni, Mubashir Noman 외 arxiv

Multimodal speaker identification systems typically assume the availability of complete and homogeneous audio-visual modalities during both training and testing. However, in real-world applications, such assumptions ofte…

Speaker Identification

Tackling the Score Shift in Cross-Lingual Speaker Verification by Exploiting Language Information

2021-10-18 · Jenthe Thienpondt, Brecht Desplanques, Kris Demuynck

This paper contains a post-challenge performance analysis on cross-lingual speaker verification of the IDLab submission to the VoxCeleb Speaker Recognition Challenge 2021 (VoxSRC-21). We show that current speaker embeddi…

Language IdentificationSpeaker RecognitionSpeaker Verification

Weakly Supervised Training of Hierarchical Attention Networks for Speaker Identification

2020-05-15 · Yanpei Shi, Qiang Huang, Thomas Hain

Identifying multiple speakers without knowing where a speaker's voice is in a recording is a challenging task. In this paper, a hierarchical attention network is proposed to solve a weakly labelled speaker identification…

Speaker Identification