On the relevance of bandwidth extension for speaker identification
In this paper we discuss the relevance of bandwidth extension for speaker identification tasks. Mainly we want to study if it is possible to recognize voices that have been bandwith extended. For this purpose, we created two different databases (microphonic and ISDN) of speech signals that were bandwidth extended from telephone bandwidth ([300, 3400] Hz) to full bandwidth ([100, 8000] Hz). We have evaluated different parameterizations, and we have found that the MELCEPST parameterization can take advantage of the bandwidth extension algorithms in several situations.
Code (0)
등록된 구현이 없습니다.
Tasks
Bandwidth ExtensionSpeaker IdentificationSimilar Papers 제목 키워드 기반
French Listening Tests for the Assessment of Intelligibility, Quality, and Identity of Body-Conducted Speech Enhancement
This study evaluates the Extreme Bandwidth Extension Network (EBEN) model on body-conduction sensors through listening tests. Using the Vibravox dataset, we assess intelligibility with a French Modified Rhyme Test, speec…
Bandwidth ExtensionSpeaker IdentificationSpeaker VerificationSpeech EnhancementA Unified Deep Speaker Embedding Framework for Mixed-Bandwidth Speech Data
This paper proposes a unified deep speaker embedding framework for modeling speech data with different sampling rates. Considering the narrowband spectrogram as a sub-image of the wideband spectrogram, we tackle the join…
Bandwidth Extensionimage-classificationImage ClassificationJoint domain adaptation and speech bandwidth extension using time-domain GANs for speaker verification
Speech systems developed for a particular choice of acoustic domain and sampling frequency do not translate easily to others. The usual practice is to learn domain adaptation and bandwidth extension models independently.…
Bandwidth ExtensionDomain AdaptationSpeaker VerificationSelf-FiLM: Conditioning GANs with self-supervised representations for bandwidth extension based speaker recognition
Speech super-resolution/Bandwidth Extension (BWE) can improve downstream tasks like Automatic Speaker Verification (ASV). We introduce a simple novel technique called Self-FiLM to inject self-supervision into existing BW…
Bandwidth ExtensionSpeaker RecognitionSpeaker VerificationSuper-Resolution+1Time-domain speech super-resolution with GAN based modeling for telephony speaker verification
Automatic Speaker Verification (ASV) technology has become commonplace in virtual assistants. However, its performance suffers when there is a mismatch between the train and test domains. Mixed bandwidth training, i.e., …
Bandwidth ExtensionData AugmentationGenerative Adversarial NetworkSpeaker Verification+1