paper-with-me

Papers

VoxCeleb2: Deep Speaker Recognition

2018-06-14 · Joon Son Chung, Arsha Nagrani, Andrew Zisserman

The objective of this paper is speaker recognition under noisy and unconstrained conditions. We make two key contributions. First, we introduce a very large-scale audio-visual speaker recognition dataset collected from open-source media. Using a fully automated pipeline, we curate VoxCeleb2 which contains over a million utterances from over 6,000 speakers. This is several times larger than any publicly available speaker recognition dataset. Second, we develop and compare Convolutional Neural Network (CNN) models and training strategies that can effectively recognise identities from voice under various conditions. The models trained on the VoxCeleb2 dataset surpass the performance of previous works on a benchmark dataset by a significant margin.

📄 PDF Abstract BibTeX arXiv:1806.05622

Code (2)

MainRo/deep-speaker
a-nagrani/VGGVox

Tasks

Speaker RecognitionSpeaker Verification

Similar Papers 제목 키워드 기반

Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices

2023-12-20 · Beltrán Labrador, Manuel Otero-Gonzalez, Alicia Lozano-Diez, Daniel Ramos 외

This paper presents VoxCeleb-ESP, a collection of pointers and timestamps to YouTube videos facilitating the creation of a novel speaker recognition dataset. VoxCeleb-ESP captures real-world scenarios, incorporating dive…

Speaker IdentificationSpeaker Recognition

VoxSRC 2019: The first VoxCeleb Speaker Recognition Challenge

2019-12-05 · Joon Son Chung, Arsha Nagrani, Ernesto Coto, Weidi Xie 외

The VoxCeleb Speaker Recognition Challenge 2019 aimed to assess how well current speaker recognition technology is able to identify speakers in unconstrained or `in the wild' data. It consisted of: (i) a publicly availab…

Speaker Recognition

ChinaTelecom System Description to VoxCeleb Speaker Recognition Challenge 2023

2023-08-16 · Mengjie Du, Xiang Fang, Jie Li

This technical report describes ChinaTelecom system for Track 1 (closed) of the VoxCeleb2023 Speaker Recognition Challenge (VoxSRC 2023). Our system consists of several ResNet variants trained only on VoxCeleb2, which we…

Speaker Recognition

The ID R&D VoxCeleb Speaker Recognition Challenge 2023 System Description

2023-08-16 · Nikita Torgashov, Rostislav Makarov, Ivan Yakovlev, Pavel Malov 외

This report describes ID R&D team submissions for Track 2 (open) to the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23). Our solution is based on the fusion of deep ResNets and self-supervised learning (SSL) base…

Self-Supervised LearningSpeaker Recognition

ShaneRun System Description to VoxCeleb Speaker Recognition Challenge 2020

2020-11-03 · Shen Chen

In this report, we describe the submission of ShaneRun's team to the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020. We use ResNet-34 as encoder to extract the speaker embeddings, which is referenced from the open-…

Speaker Recognition