paper-with-me

Papers

Learnable MFCCs for Speaker Verification

2021-02-20 · Xuechen Liu, Md Sahidullah, Tomi Kinnunen

We propose a learnable mel-frequency cepstral coefficient (MFCC) frontend architecture for deep neural network (DNN) based automatic speaker verification. Our architecture retains the simplicity and interpretability of MFCC-based features while allowing the model to be adapted to data flexibly. In practice, we formulate data-driven versions of the four linear transforms of a standard MFCC extractor -- windowing, discrete Fourier transform (DFT), mel filterbank and discrete cosine transform (DCT). Results reported reach up to 6.7\% (VoxCeleb1) and 9.7\% (SITW) relative improvement in term of equal error rate (EER) from static MFCCs, without additional tuning effort.

📄 PDF Abstract BibTeX arXiv:2102.10322

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Verification

Methods 이 논문이 사용한 방법론

Discrete Cosine Transform Discrete Cosine Transform (DCT) is an orthogonal transformation method that decomposes an image to its spatial frequency spectrum. It expresses a finite sequence of data…

Similar Papers 제목 키워드 기반

Optimizing Multi-Taper Features for Deep Speaker Verification

2021-10-21 · Xuechen Liu, Md Sahidullah, Tomi Kinnunen

Multi-taper estimators provide low-variance power spectrum estimates that can be used in place of the windowed discrete Fourier transform (DFT) to extract speech features such as mel-frequency cepstral coefficients (MFCC…

Open-Ended Question AnsweringSpeaker Verification

A Comparison of Features for Replay Attack Detection

2019-12-12 · IOP Conf. Series: Journal of Physics: Conf. Series 1229 2019 12 · Zhifeng Xiea, Weibin Zhangb, Zhuxin Chen and Xiangmin Xu

Speaker verification (ASV) systems are still vulnerable to different kinds of spoofing attacks, especially replay attack due to high-quality playback devices. Many countermeasures have been developed recently. Most of th…

Speaker Verification

Robust Support Vector Machines for Speaker Verification Task

2013-06-12 · Kawthar Yasmine Zergat, Abderrahmane Amrouche

An important step in speaker verification is extracting features that best characterize the speaker voice. This paper investigates a front-end processing that aims at improving the performance of speaker verification bas…

Dimensionality ReductionSpeaker Verification

Optimization of data-driven filterbank for automatic speaker verification

2020-07-21 · Susanta Sarangi, Md Sahidullah, Goutam Saha

Most of the speech processing applications use triangular filters spaced in mel-scale for feature extraction. In this paper, we propose a new data-driven filter design method which optimizes filter parameters from a give…

Speaker Verification

Histogram Transform-based Speaker Identification

2018-08-02 · Zhanyu Ma, Hong Yu

A novel text-independent speaker identification (SI) method is proposed. This method uses the Mel-frequency Cepstral coefficients (MFCCs) and the dynamic information among adjacent frames as feature sets to capture speak…

Speaker Identification