paper-with-me

Papers

Deep Multi-Frame MVDR Filtering for Single-Microphone Speech Enhancement

2020-11-20 · Marvin Tammen, Simon Doclo

Multi-frame algorithms for single-microphone speech enhancement, e.g., the multi-frame minimum variance distortionless response (MFMVDR) filter, are able to exploit speech correlation across adjacent time frames in the short-time Fourier transform (STFT) domain. Provided that accurate estimates of the required speech interframe correlation vector and the noise correlation matrix are available, it has been shown that the MFMVDR filter yields a substantial noise reduction while hardly introducing any speech distortion. Aiming at merging the speech enhancement potential of the MFMVDR filter and the estimation capability of temporal convolutional networks (TCNs), in this paper we propose to embed the MFMVDR filter within a deep learning framework. The TCNs are trained to map the noisy speech STFT coefficients to the required quantities by minimizing the scale-invariant signal-to-distortion ratio loss function at the MFMVDR filter output. Experimental results show that the proposed deep MFMVDR filter achieves a competitive speech enhancement performance on the Deep Noise Suppression Challenge dataset. In particular, the results show that estimating the parameters of an MFMVDR filter yields a higher performance in terms of PESQ and STOI than directly estimating the multi-frame filter or single-frame masks and than Conv-TasNet.

📄 PDF Abstract BibTeX arXiv:2011.10345

Code (1)

https://gitlab.uni-oldenburg.de/hura4843/deep-mfmvdr 공식 구현

Tasks

Speech Enhancement

Similar Papers 제목 키워드 기반

Deep Multi-Frame MVDR Filtering for Binaural Noise Reduction

2022-05-18 · Marvin Tammen, Simon Doclo

To improve speech intelligibility and speech quality in noisy environments, binaural noise reduction algorithms for head-mounted assistive listening devices are of crucial importance. Several binaural noise reduction alg…

RTF-steered binaural MVDR beamforming incorporating multiple external microphones

2019-08-13

The binaural minimum-variance distortionless-response (BMVDR) beamformer is a well-known noise reduction algorithm that can be steered using the relative transfer function (RTF) vector of the desired speech source. Explo…

Multi-microphone Complex Spectral Mapping for Utterance-wise and Continuous Speech Separation

2020-10-04 · Zhong-Qiu Wang, Peidong Wang, DeLiang Wang

We propose multi-microphone complex spectral mapping, a simple way of applying deep learning for time-varying non-linear beamforming, for speaker separation in reverberant conditions. We aim at both speaker separation an…

Speaker SeparationSpeech Separation

Deep Multi-Frame Filtering for Hearing Aids

2023-05-14 · Hendrik Schröter, Tobias Rosenkranz, Alberto N. Escalante-B., Andreas Maier

Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep filtering (DF) recently demonstrated its capabilities for low-latency sc…

Speech Enhancement

ChannelAugment: Improving generalization of multi-channel ASR by training with input channel randomization

2021-09-23 · Marco Gaudesi, Felix Weninger, Dushyant Sharma, Puming Zhan

End-to-end (E2E) multi-channel ASR systems show state-of-the-art performance in far-field ASR tasks by joint training of a multi-channel front-end along with the ASR model. The main limitation of such systems is that the…

Data Augmentation