paper-with-me

홈 › Papers

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin

2025-05-27 · Zhiqi Ai, Meixuan Bao, Zhiyong Chen, Zhi Yang, Xinnuo Li, Shugong Xu

The performance of speaker verification systems is adversely affected by speaker aging. However, due to challenges in data collection, particularly the lack of sustained and large-scale longitudinal data for individuals, research on speaker aging remains difficult. In this paper, we present VoxAging, a large-scale longitudinal dataset collected from 293 speakers (226 English speakers and 67 Mandarin speakers) over several years, with the longest time span reaching 17 years (approximately 900 weeks). For each speaker, the data were recorded at weekly intervals. We studied the phenomenon of speaker aging and its effects on advanced speaker verification systems, analyzed individual speaker aging processes, and explored the impact of factors such as age group and gender on speaker aging research.

📄 PDF Abstract BibTeX arXiv:2505.21445

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Similar Papers 제목 키워드 기반

Towards Low-Latency Tracking of Multiple Speakers With Short-Context Speaker Embeddings

2025-08-18 · Taous Iatariene, Alexandre Guérin, Romain Serizel arxiv

Speaker embeddings are promising identity-related features that can enhance the identity assignment performance of a tracking system by leveraging its spatial predictions, i.e, by performing identity reassignment. Common…

Knowledge Distillation

Autoregressive Guidance of Deep Spatially Selective Filters using Bayesian Tracking for Efficient Extraction of Moving Speakers

2026-03-24 · Jakob Kienegger, Timo Gerkmann arxiv

Deep spatially selective filters achieve high-quality enhancement with real-time capable architectures for stationary speakers of known directions. To retain this level of performance in dynamic scenarios where only the …

Speech Trax: A Bottom to the Top Approach for Speaker Tracking and Indexing in an Archiving Context

2016-05-01 · LREC 2016 5 · F{\'e}licien Vallet, Jim Uro, J{\'e}r{\'e}my Andriamakaoly, Hakim Nabi 외

With the increasing amount of audiovisual and digital data deriving from televisual and radiophonic sources, professional archives such as INA, France{'}s national audiovisual institute, acknowledge a growing need for ef…

speaker-diarizationSpeaker Diarization

Speaker Embeddings to Improve Tracking of Intermittent and Moving Speakers

2025-06-23 · Taous Iatariene, Can Cui, Alexandre Guérin, Romain Serizel

Speaker tracking methods often rely on spatial observations to assign coherent track identities over time. This raises limits in scenarios with intermittent and moving speakers, i.e., speakers that may change position wh…

Position

Jointly Tracking and Separating Speech Sources Using Multiple Features and the generalized labeled multi-Bernoulli Framework

2018-04-16

This paper proposes a novel joint multi-speaker tracking-and-separation method based on the generalized labeled multi-Bernoulli (GLMB) multi-target tracking filter, using sound mixtures recorded by microphones. Standard …