paper-with-me

Papers

A Review of Multi-Objective Deep Learning Speech Denoising Methods

2020-03-26

This paper presents a review of multi-objective deep learning methods that have been introduced in the literature for speech denoising. After stating an overview of conventional, single objective deep learning, and hybrid or combined conventional and deep learning methods, a review of the mathematical framework of the multi-objective deep learning methods for speech denoising is provided. A representative method from each speech denoising category, whose codes are publicly available, is selected and a comparison is carried out by considering the same public domain dataset and four widely used objective metrics. The comparison results indicate the effectiveness of the multi-objective method compared with the other methods, in particular when the signal-to-noise ratio is low. Possible future improvements that can be achieved are also mentioned.

📄 PDF Abstract BibTeX arXiv:2003.12108

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningDenoisingSpeech Denoising

Similar Papers 제목 키워드 기반

LuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish

2026-06-30 · Nina Hosseini-Kivanani, Sandipana Dowerah arxiv

State-of-the-art speech datasets predominantly focus on widely spoken languages, often overlooking low-resource languages such as Luxembourgish, which remain underrepresented in speech technology research. In this work, …

Language IdentificationCross-Lingual TransferActivity Detection

CleanUNet 2: A Hybrid Speech Denoising Model on Waveform and Spectrogram

2023-09-12 · Zhifeng Kong, Wei Ping, Ambrish Dantrey, Bryan Catanzaro

In this work, we present CleanUNet 2, a speech denoising model that combines the advantages of waveform denoiser and spectrogram denoiser and achieves the best of both worlds. CleanUNet 2 uses a two-stage framework inspi…

DenoisingSpeech DenoisingSpeech EnhancementSpeech Synthesis

Speech Denoising in the Waveform Domain with Self-Attention

2022-02-15 · Zhifeng Kong, Wei Ping, Ambrish Dantrey, Bryan Catanzaro

In this work, we present CleanUNet, a causal speech denoising model on the raw waveform. The proposed model is based on an encoder-decoder architecture combined with several self-attention blocks to refine its bottleneck…

DecoderDenoisingSpeech DenoisingSpeech Enhancement

Continuous Modeling of the Denoising Process for Speech Enhancement Based on Deep Learning

2023-09-17 · Zilu Guo, Jun Du, Chin-Hui Lee

In this paper, we explore a continuous modeling approach for deep-learning-based speech enhancement, focusing on the denoising process. We use a state variable to indicate the denoising process. The starting state is noi…

Automatic Speech RecognitionDenoisingSpeech Enhancementspeech-recognition+1

Zero-Shot Voice Conditioning for Denoising Diffusion TTS Models

2022-06-05 · Alon Levkovitch, Eliya Nachmani, Lior Wolf

We present a novel way of conditioning a pretrained denoising diffusion speech model to produce speech in the voice of a novel person unseen during training. The method requires a short (~3 seconds) sample from the targe…

Denoising