paper-with-me

홈 › Papers

Speech Denoising Convolutional Neural Network trained with Deep Feature Losses.

2018-06-27 · Interspeech 2018 6 · Francois G. Germain, Qifeng Chen, Vladlen Koltun

We present an end-to-end deep learning approach to denoising speech signals by processing the raw waveform directly. Given input audio containing speech corrupted by an additive background signal, the system aims to produce a processed signal that contains only the speech content. Recent approaches have shown promising results using various deep network architectures. In this paper, we propose to train a fully-convolutional context aggregation network using a deep feature loss. That loss is based on comparing the internal feature activations in a different network, trained for acoustic environment detection and domestic audio tagging. Our approach outperforms the state-of-the-art in objective speech quality metrics and in large-scale perceptual experiments with human listeners. It also outperforms an identical network trained using traditional regression losses. The advantage of the new approach is particularly pronounced for the hardest data with the most intrusive background noise, for which denoising is most needed and most challenging.

📄 PDF Abstract BibTeX

Code (2)

0three/Speech-Denoise-With-Feature-Loss
francoisgermain/SpeechDenoisingWithDeepFeatureLosses tf

Tasks

Audio TaggingDenoisingSpeech DenoisingSpeech Enhancement

Similar Papers 제목 키워드 기반

Speech Denoising with Deep Feature Losses

2018-06-27 · Francois G. Germain, Qifeng Chen, Vladlen Koltun

We present an end-to-end deep learning approach to denoising speech signals by processing the raw waveform directly. Given input audio containing speech corrupted by an additive background signal, the system aims to prod…

Audio TaggingDenoisingSpeech Denoising

Speech Denoising with Auditory Models

2020-11-21 · Mark R. Saddler, Andrew Francl, Jenelle Feather, Kaizhi Qian 외

Contemporary speech enhancement predominantly relies on audio transforms that are trained to reconstruct a clean speech waveform. The development of high-performing neural network sound recognition systems has raised the…

DenoisingSpeech DenoisingSpeech Enhancement

HiFi-GAN: High-Fidelity Denoising and Dereverberation Based on Speech Deep Features in Adversarial Networks

2020-06-10 · Jiaqi Su, Zeyu Jin, Adam Finkelstein

Real-world audio recordings are often degraded by factors such as noise, reverberation, and equalization distortion. This paper introduces HiFi-GAN, a deep learning method to transform recorded speech to sound as though …

DenoisingSpeech DereverberationSpeech Enhancement

Deep speech inpainting of time-frequency masks

2019-10-20 · Mikolaj Kegler, Pierre Beckmann, Milos Cernak

Transient loud intrusions, often occurring in noisy environments, can completely overpower speech signal and lead to an inevitable loss of information. While existing algorithms for noise suppression can yield impressive…

Retrieval

Perceptual Loss based Speech Denoising with an ensemble of Audio Pattern Recognition and Self-Supervised Models

2020-10-22

Deep learning based speech denoising still suffers from the challenge of improving perceptual quality of enhanced signals. We introduce a generalized framework called Perceptual Ensemble Regularization Loss (PERL) built …

DenoisingEmotion ClassificationMulti-Task LearningSpeech Denoising