paper-with-me

Papers

WPD++: An Improved Neural Beamformer for Simultaneous Speech Separation and Dereverberation

2020-11-18

This paper aims at eliminating the interfering speakers' speech, additive noise, and reverberation from the noisy multi-talker speech mixture that benefits automatic speech recognition (ASR) backend. While the recently proposed Weighted Power minimization Distortionless response (WPD) beamformer can perform separation and dereverberation simultaneously, the noise cancellation component still has the potential to progress. We propose an improved neural WPD beamformer called "WPD++" by an enhanced beamforming module in the conventional WPD and a multi-objective loss function for the joint training. The beamforming module is improved by utilizing the spatio-temporal correlation. A multi-objective loss, including the complex spectra domain scale-invariant signal-to-noise ratio (C-Si-SNR) and the magnitude domain mean square error (Mag-MSE), is properly designed to make multiple constraints on the enhanced speech and the desired power of the dry clean signal. Joint training is conducted to optimize the complex-valued mask estimator and the WPD++ beamformer in an end-to-end way. The results show that the proposed WPD++ outperforms several state-of-the-art beamformers on the enhanced speech quality and word error rate (WER) of ASR.

📄 PDF Abstract BibTeX arXiv:2011.09162

Code (1)

nateanl/wpd-plus-plus

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionSpeech Separation

Similar Papers 제목 키워드 기반

End-to-End Far-Field Speech Recognition with Unified Dereverberation and Beamforming

2020-10-27

Despite successful applications of end-to-end approaches in multi-channel speech recognition, the performance still degrades severely when the speech is corrupted by reverberation. In this paper, we integrate the derever…

speech-recognitionSpeech Recognition

An Effective Dereverberation Algorithm by Fusing MVDR and MCLP

2022-03-28 · Fengqi Tan, Changchun Bao

In the scenario with reverberation, the experience of human-machine interaction will become worse. In order to solve this problem, many methods for the dereverberation have emerged. At present, how to update the paramete…

Blocking

Jointly optimal dereverberation and beamforming

2019-10-30 · Christoph Boeddeker, Tomohiro Nakatani, Keisuke Kinoshita, Reinhold Haeb-Umbach

We previously proposed an optimal (in the maximum likelihood sense) convolutional beamformer that can perform simultaneous denoising and dereverberation, and showed its superiority over the widely used cascade of a WPE d…

Denoising

Blind and neural network-guided convolutional beamformer for joint denoising, dereverberation, and source separation

2021-08-04 · Tomohiro Nakatani, Rintaro Ikeshita, Keisuke Kinoshita, Hiroshi Sawada 외

This paper proposes an approach for optimizing a Convolutional BeamFormer (CBF) that can jointly perform denoising (DN), dereverberation (DR), and source separation (SS). First, we develop a blind CBF optimization algori…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Denoisingspeech-recognition+1

Blind Speech Separation and Dereverberation using Neural Beamforming

2021-03-24 · Lukas Pfeifenberger, Franz Pernkopf

In this paper, we present the Blind Speech Separation and Dereverberation (BSSD) network, which performs simultaneous speaker separation, dereverberation and speaker identification in a single neural network. Speaker sep…

Speaker IdentificationSpeaker SeparationSpeech SeparationTriplet