paper-with-me

홈 › Papers

Large-scale Speaker Retrieval on Random Speaker Variability Subspace

2018-11-27 · Suwon Shon, Young-Gun Lee, Taesu Kim

This paper describes a fast speaker search system to retrieve segments of the same voice identity in the large-scale data. A recent study shows that Locality Sensitive Hashing (LSH) enables quick retrieval of a relevant voice in the large-scale data in conjunction with i-vector while maintaining accuracy. In this paper, we proposed Random Speaker-variability Subspace (RSS) projection to map a data into LSH based hash tables. We hypothesized that rather than projecting on completely random subspace without considering data, projecting on randomly generated speaker variability space would give more chance to put the same speaker representation into the same hash bins, so we can use less number of hash tables. Multiple RSS can be generated by randomly selecting a subset of speakers from a large speaker cohort. From the experimental result, the proposed approach shows 100 times and 7 times faster than the linear search and LSH, respectively

📄 PDF Abstract BibTeX arXiv:1811.10812

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval

Similar Papers 제목 키워드 기반

Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness

2025-04-26 · Erfan Loweimi, Mengjie Qian, Kate Knill, Mark Gales

There is a growing abundance of publicly available or company-owned audio/video archives, highlighting the increasing importance of efficient access to desired content and information retrieval from these archives. This …

Information RetrievalRetrieval

VieSpeaker: A Large-Scale Vietnamese Speaker Recognition Dataset Beyond Visual Dependency

2026-06-23 · Viet Hoang Pham, Tran Trung Nguyen, Bao Thu Ho, Phuong Tuan Dat 외 arxiv

Speaker recognition has advanced rapidly with large-scale training datasets, yet Vietnamese remains under-resourced, with existing corpora limited in scale and acoustic diversity. Most large-scale datasets rely on facial…

Speaker Recognition

Bilingual Text-dependent Speaker Verification with Pre-trained Models for TdSV Challenge 2024

2024-11-16 · Seyed Ali Farokh

This paper presents our submissions to the Iranian division of the Text-dependent Speaker Verification Challenge (TdSV) 2024. TdSV aims to determine if a specific phrase was spoken by a target speaker. We developed two i…

Domain AdaptationSpeaker VerificationText-Dependent Speaker Verification

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin

2025-05-27 · Zhiqi Ai, Meixuan Bao, Zhiyong Chen, Zhi Yang 외

The performance of speaker verification systems is adversely affected by speaker aging. However, due to challenges in data collection, particularly the lack of sustained and large-scale longitudinal data for individuals,…

Speaker Verification

Large-Scale Speaker Diarization of Radio Broadcast Archives

2019-06-19 · Emre Yilmaz, Adem Derinel, Zhou Kun, Henk van den Heuvel 외

This paper describes our initial efforts to build a large-scale speaker diarization (SD) and identification system on a recently digitized radio broadcast archive from the Netherlands which has more than 6500 audio tapes…

speaker-diarizationSpeaker DiarizationSpeaker Identification