paper-with-me

홈 › Papers

Speaker recognition improvement using blind inversion of distortions

2022-02-23 · Marcos Faundez-Zanuy, Jordi Sole-Casals

In this paper we propose the inversion of nonlinear distortions in order to improve the recognition rates of a speaker recognizer system. We study the effect of saturations on the test signals, trying to take into account real situations where the training material has been recorded in a controlled situation but the testing signals present some mismatch with the input signal level (saturations). The experimental results shows that a combination of data fusion with and without nonlinear distortion compensation can improve the recognition rates with saturated test sentences from 80% to 88.57%, while the results with clean speech (without saturation) is 87.76% for one microphone.

📄 PDF Abstract BibTeX arXiv:2203.01164

Code (0)

등록된 구현이 없습니다.

Tasks

Speaker Recognition

Similar Papers 제목 키워드 기반

Joint Sound Source Separation and Speaker Recognition

2016-04-29 · Jeroen Zegers, Hugo Van hamme

Non-negative Matrix Factorization (NMF) has already been applied to learn speaker characterizations from single or non-simultaneous speech for speaker recognition applications. It is also known for its good performance i…

blind source separationSpeaker Recognition

Introducing Model Inversion Attacks on Automatic Speaker Recognition

2023-01-09 · Karla Pizzi, Franziska Boenisch, Ugur Sahin, Konstantin Böttinger

Model inversion (MI) attacks allow to reconstruct average per-class representations of a machine learning (ML) model's training data. It has been shown that in scenarios where each class corresponds to a different indivi…

modelSpeaker Recognition

Articulatory-WaveNet: Autoregressive Model For Acoustic-to-Articulatory Inversion

2020-06-22 · Narjes Bozorg, Michael T. Johnson

This paper presents Articulatory-WaveNet, a new approach for acoustic-to-articulator inversion. The proposed system uses the WaveNet speech synthesis architecture, with dilated causal convolutional layers using previous …

Speech Synthesis

Blindly Assess Quality of In-the-Wild Videos via Quality-aware Pre-training and Motion Perception

2021-08-19 · Bowen Li, Weixia Zhang, Meng Tian, Guangtao Zhai 외

Perceptual quality assessment of the videos acquired in the wilds is of vital importance for quality assurance of video services. The inaccessibility of reference videos with pristine quality and the complexity of authen…

Action RecognitionImage Quality AssessmentTransfer LearningVideo Quality Assessment+1

Blind score normalization method for PLDA based speaker recognition

2016-02-22 · Danila Doroshin, Nikolay Lubimov, Marina Nastasenko, Mikhail Kotov

Probabilistic Linear Discriminant Analysis (PLDA) has become state-of-the-art method for modeling $i$-vector space in speaker recognition task. However the performance degradation is observed if enrollment data size diff…

Speaker Recognition