paper-with-me

홈 › Papers

Enhancing Open-Set Speaker Identification through Rapid Tuning with Speaker Reciprocal Points and Negative Sample

2024-09-24 · Zhiyong Chen, Zhiqi Ai, Xinnuo Li, Shugong Xu

This paper introduces a novel framework for open-set speaker identification in household environments, playing a crucial role in facilitating seamless human-computer interactions. Addressing the limitations of current speaker models and classification approaches, our work integrates an pretrained WavLM frontend with a few-shot rapid tuning neural network (NN) backend for enrollment, employing task-optimized Speaker Reciprocal Points Learning (SRPL) to enhance discrimination across multiple target speakers. Furthermore, we propose an enhanced version of SRPL (SRPL+), which incorporates negative sample learning with both speech-synthesized and real negative samples to significantly improve open-set SID accuracy. Our approach is thoroughly evaluated across various multi-language text-dependent speaker recognition datasets, demonstrating its effectiveness in achieving high usability for complex household multi-speaker recognition scenarios. The proposed system enhanced open-set performance by up to 27\% over the directly use of efficient WavLM base+ model.

📄 PDF Abstract BibTeX arXiv:2409.15742

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker IdentificationSpeaker Recognition

Similar Papers 제목 키워드 기반

Speaker identification from the sound of the human breath

2017-12-01 · Wenbo Zhao, Yang Gao, Rita Singh

This paper examines the speaker identification potential of breath sounds in continuous speech. Speech is largely produced during exhalation. In order to replenish air in the lungs, speakers must periodically inhale. Whe…

Speaker IdentificationSpeaker Recognition

openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer

2022-02-24 · Kishan K C, Zhenning Tan, Long Chen, Minho Jin 외

Household speaker identification with few enrollment utterances is an important yet challenging problem, especially when household members share similar voice characteristics and room acoustics. A common embedding space …

Open Set LearningSpeaker Identification

Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models

2024-07-16 · Minh Nguyen, Franck Dernoncourt, Seunghyun Yoon, Hanieh Deilamsalehy 외

We introduce an approach to identifying speaker names in dialogue transcripts, a crucial task for enhancing content accessibility and searchability in digital media archives. Despite the advancements in speech recognitio…

AttributeSpeaker Identificationspeech-recognitionSpeech Recognition

Rapid IoT Device Identification at the Edge

2021-10-26 · Oliver Thompson, Anna Maria Mandalari, Hamed Haddadi

Consumer Internet of Things (IoT) devices are increasingly common in everyday homes, from smart speakers to security cameras. Along with their benefits come potential privacy and security threats. To limit these threats …

IoT Device Identification

Experiments on Open-Set Speaker Identification with Discriminatively Trained Neural Networks

2019-04-02 · Stefano Imoscopi, Volodya Grancharov, Sigurdur Sverrisson, Erlendur Karlsson 외

This paper presents a study on discriminative artificial neural network classifiers in the context of open-set speaker identification. Both 2-class and multi-class architectures are tested against the conventional Gaussi…

Speaker Identification