paper-with-me

홈 › Papers

SkipConvGAN: Monaural Speech Dereverberation using Generative Adversarial Networks via Complex Time-Frequency Masking

2022-11-22 · Vinay Kothapally, J. H. L. Hansen

With the advancements in deep learning approaches, the performance of speech enhancing systems in the presence of background noise have shown significant improvements. However, improving the system's robustness against reverberation is still a work in progress, as reverberation tends to cause loss of formant structure due to smearing effects in time and frequency. A wide range of deep learning-based systems either enhance the magnitude response and reuse the distorted phase or enhance complex spectrogram using a complex time-frequency mask. Though these approaches have demonstrated satisfactory performance, they do not directly address the lost formant structure caused by reverberation. We believe that retrieving the formant structure can help improve the efficiency of existing systems. In this study, we propose SkipConvGAN - an extension of our prior work SkipConvNet. The proposed system's generator network tries to estimate an efficient complex time-frequency mask, while the discriminator network aids in driving the generator to restore the lost formant structure. We evaluate the performance of our proposed system on simulated and real recordings of reverberant speech from the single-channel task of the REVERB challenge corpus. The proposed system shows a consistent improvement across multiple room configurations over other deep learning-based generative adversarial frameworks.

📄 PDF Abstract BibTeX arXiv:2211.12623

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningSpeech Dereverberation

Similar Papers 제목 키워드 기반

CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement

2022-09-22 · Sherif Abdulatif, Ruizhe Cao, Bin Yang

In this work, we further develop the conformer-based metric generative adversarial network (CMGAN) model for speech enhancement (SE) in the time-frequency (TF) domain. This paper builds on our previous work but takes a m…

Audio Super-ResolutionAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Decoder+8

Receptive Field Analysis of Temporal Convolutional Networks for Monaural Speech Dereverberation

2022-04-13 · William Ravenscroft, Stefan Goetze, Thomas Hain

Speech dereverberation is often an important requirement in robust speech processing tasks. Supervised deep learning (DL) models give state-of-the-art performance for single-channel speech dereverberation. Temporal convo…

Speech DereverberationSpeech Enhancement

Investigating Generative Adversarial Networks based Speech Dereverberation for Robust Speech Recognition

2018-03-27 · Ke Wang, Junbo Zhang, Sining Sun, Yujun Wang 외

We investigate the use of generative adversarial networks (GANs) in speech dereverberation for robust speech recognition. GANs have been recently studied for speech enhancement to remove additive noises, but there still …

Robust Speech RecognitionSpeech DereverberationSpeech Enhancementspeech-recognition+1

Simultaneous Denoising and Dereverberation Using Deep Embedding Features

2020-04-06 · Cunhang Fan, Jian-Hua Tao, Bin Liu, Jiangyan Yi 외

Monaural speech dereverberation is a very challenging task because no spatial cues can be used. When the additive noises exist, this task becomes more challenging. In this paper, we propose a joint training method for si…

ClusteringDeep ClusteringDenoisingSpeech Denoising+2

Single-channel Speech Dereverberation via Generative Adversarial Training

2018-06-25 · Chenxing Li, Tieqiang Wang, Shuang Xu, Bo Xu

In this paper, we propose a single-channel speech dereverberation system (DeReGAT) based on convolutional, bidirectional long short-term memory and deep feed-forward neural network (CBLDNN) with generative adversarial tr…

Speech Dereverberation