paper-with-me

홈 › Papers

Deep learning methods in speaker recognition: a review

2019-11-14 · Dávid Sztahó, György Szaszák, András Beke

This paper summarizes the applied deep learning practices in the field of speaker recognition, both verification and identification. Speaker recognition has been a widely used field topic of speech technology. Many research works have been carried out and little progress has been achieved in the past 5-6 years. However, as deep learning techniques do advance in most machine learning fields, the former state-of-the-art methods are getting replaced by them in speaker recognition too. It seems that DL becomes the now state-of-the-art solution for both speaker verification and identification. The standard x-vectors, additional to i-vectors, are used as baseline in most of the novel works. The increasing amount of gathered data opens up the territory to DL, where they are the most effective.

📄 PDF Abstract BibTeX arXiv:1911.06615

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningSpeaker RecognitionSpeaker Verification

Similar Papers 제목 키워드 기반

Speaker Recognition Based on Deep Learning: An Overview

2020-12-02 · Zhongxin Bai, Xiao-Lei Zhang

Speaker recognition is a task of identifying persons from their voices. Recently, deep learning has dramatically revolutionized speaker recognition. However, there is lack of comprehensive reviews on the exciting progres…

Deep LearningDomain Adaptationspeaker-diarizationSpeaker Diarization+3

A Review of Speaker Diarization: Recent Advances with Deep Learning

2021-01-24 · Tae Jin Park, Naoyuki Kanda, Dimitrios Dimitriadis, Kyu J. Han 외

Speaker diarization is a task to label audio or video recordings with classes that correspond to speaker identity, or in short, a task to identify "who spoke when". In the early years, speaker diarization algorithms were…

Deep LearningRetrievalspeaker-diarizationSpeaker Diarization+2

Utterance partitioning for speaker recognition: an experimental review and analysis with new findings under GMM-SVM framework

2021-05-25 · Nirmalya Sen, Md Sahidullah, Hemant Patil, Shyamal Kumar Das Mandal 외

The performance of speaker recognition system is highly dependent on the amount of speech used in enrollment and test. This work presents a detailed experimental review and analysis of the GMM-SVM based speaker recogniti…

Speaker Recognition

CNVSRC 2023: The First Chinese Continuous Visual Speech Recognition Challenge

2024-06-14 · Chen Chen, Zehua Liu, Xiaolou Li, Lantian Li 외

The first Chinese Continuous Visual Speech Recognition Challenge aimed to probe the performance of Large Vocabulary Continuous Visual Speech Recognition (LVC-VSR) on two tasks: (1) Single-speaker VSR for a particular spe…

speech-recognitionSpeech RecognitionVisual Speech Recognition

A Comprehensive Survey on Multi-modal Conversational Emotion Recognition with Deep Learning

2023-12-10 · Yuntao Shou, Tao Meng, Wei Ai, Nan Yin 외

Multi-modal conversation emotion recognition (MCER) aims to recognize and track the speaker's emotional state using text, speech, and visual information in the conversation scene. Analyzing and studying MCER issues is si…

Emotion Recognition