paper-with-me

홈 › Papers

Improving Speaker Diarization using Semantic Information: Joint Pairwise Constraints Propagation

2023-09-19 · Luyao Cheng, Siqi Zheng, Qinglin Zhang, Hui Wang, Yafeng Chen, Qian Chen, Shiliang Zhang

Speaker diarization has gained considerable attention within speech processing research community. Mainstream speaker diarization rely primarily on speakers' voice characteristics extracted from acoustic signals and often overlook the potential of semantic information. Considering the fact that speech signals can efficiently convey the content of a speech, it is of our interest to fully exploit these semantic cues utilizing language models. In this work we propose a novel approach to effectively leverage semantic information in clustering-based speaker diarization systems. Firstly, we introduce spoken language understanding modules to extract speaker-related semantic information and utilize these information to construct pairwise constraints. Secondly, we present a novel framework to integrate these constraints into the speaker diarization pipeline, enhancing the performance of the entire system. Extensive experiments conducted on the public dataset demonstrate the consistent superiority of our proposed approach over acoustic-only speaker diarization systems.

📄 PDF Abstract BibTeX arXiv:2309.10456

Code (0)

등록된 구현이 없습니다.

Tasks

speaker-diarizationSpeaker DiarizationSpoken Language Understanding

Similar Papers 제목 키워드 기반

Integrating Audio, Visual, and Semantic Information for Enhanced Multimodal Speaker Diarization

2024-08-22 · Luyao Cheng, Hui Wang, Siqi Zheng, Yafeng Chen 외

Speaker diarization, the process of segmenting an audio stream or transcribed speech content into homogenous partitions based on speaker identity, plays a crucial role in the interpretation and analysis of human speech. …

speaker-diarizationSpeaker Diarization

Exploring Speaker-Related Information in Spoken Language Understanding for Better Speaker Diarization

2023-05-22 · Luyao Cheng, Siqi Zheng, Zhang Qinglin, Hui Wang 외

Speaker diarization(SD) is a classic task in speech processing and is crucial in multi-party scenarios such as meetings and conversations. Current mainstream speaker diarization approaches consider acoustic information o…

speaker-diarizationSpeaker DiarizationSpoken Language Understanding

Enhancing Speaker Diarization with Large Language Models: A Contextual Beam Search Approach

2023-09-11 · Tae Jin Park, Kunal Dhawan, Nithin Koluguri, Jagadeesh Balam

Large language models (LLMs) have shown great promise for capturing contextual information in natural language processing tasks. We propose a novel approach to speaker diarization that incorporates the prowess of LLMs to…

speaker-diarizationSpeaker Diarization

EEND-SS: Joint End-to-End Neural Speaker Diarization and Speech Separation for Flexible Number of Speakers

2022-03-31 · Soumi Maiti, Yushi Ueda, Shinji Watanabe, Chunlei Zhang 외

In this paper, we present a novel framework that jointly performs three tasks: speaker diarization, speech separation, and speaker counting. Our proposed framework integrates speaker diarization based on end-to-end neura…

Decoderspeaker-diarizationSpeaker DiarizationSpeech Separation

Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder

2025-08-28 · Muhammad Shakeel, Yui Sudo, Yifan Peng, Chyi-Jiunn Lin 외 arxiv

This paper presents a unified multi-speaker encoder (UME), a novel architecture that jointly learns representations for speaker diarization (SD), speech separation (SS), and multi-speaker automatic speech recognition (AS…

Speaker DiarizationSpeech RecognitionSpeech Separation