paper-with-me

Papers

Multi-channel Multi-frame ADL-MVDR for Target Speech Separation

2020-12-24 · Zhuohuang Zhang, Yong Xu, Meng Yu, Shi-Xiong Zhang, LianWu Chen, Donald S. Williamson, Dong Yu

Many purely neural network based speech separation approaches have been proposed to improve objective assessment scores, but they often introduce nonlinear distortions that are harmful to modern automatic speech recognition (ASR) systems. Minimum variance distortionless response (MVDR) filters are often adopted to remove nonlinear distortions, however, conventional neural mask-based MVDR systems still result in relatively high levels of residual noise. Moreover, the matrix inverse involved in the MVDR solution is sometimes numerically unstable during joint training with neural networks. In this study, we propose a multi-channel multi-frame (MCMF) all deep learning (ADL)-MVDR approach for target speech separation, which extends our preliminary multi-channel ADL-MVDR approach. The proposed MCMF ADL-MVDR system addresses linear and nonlinear distortions. Spatio-temporal cross correlations are also fully utilized in the proposed approach. The proposed systems are evaluated using a Mandarin audio-visual corpus and are compared with several state-of-the-art approaches. Experimental results demonstrate the superiority of our proposed systems under different scenarios and across several objective evaluation metrics, including ASR performance.

📄 PDF Abstract BibTeX arXiv:2012.13442

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech RecognitionSpeech Separation

Similar Papers 제목 키워드 기반

Beam-Guided TasNet: An Iterative Speech Separation Framework with Multi-Channel Output

2021-02-05 · Hangting Chen, Yang Yi, Dang Feng, Pengyuan Zhang

Time-domain audio separation network (TasNet) has achieved remarkable performance in blind source separation (BSS). Classic multi-channel speech processing framework employs signal estimation and beamforming. For example…

blind source separationSpeech Separation

Deep Multi-Frame Filtering for Hearing Aids

2023-05-14 · Hendrik Schröter, Tobias Rosenkranz, Alberto N. Escalante-B., Andreas Maier

Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep filtering (DF) recently demonstrated its capabilities for low-latency sc…

Speech Enhancement

Subspace Hybrid Beamforming for Head-worn Microphone Arrays

2023-03-15 · Sina Hafezi, Alastair H. Moore, Pierre Guiraud, Patrick A. Naylor 외

A two-stage multi-channel speech enhancement method is proposed which consists of a novel adaptive beamformer, Hybrid Minimum Variance Distortionless Response (MVDR), Isotropic-MVDR (Iso), and a novel multi-channel spect…

DenoisingSpeech Enhancement

An Effective Dereverberation Algorithm by Fusing MVDR and MCLP

2022-03-28 · Fengqi Tan, Changchun Bao

In the scenario with reverberation, the experience of human-machine interaction will become worse. In order to solve this problem, many methods for the dereverberation have emerged. At present, how to update the paramete…

Blocking

Deep Multi-Frame MVDR Filtering for Binaural Noise Reduction

2022-05-18 · Marvin Tammen, Simon Doclo

To improve speech intelligibility and speech quality in noisy environments, binaural noise reduction algorithms for head-mounted assistive listening devices are of crucial importance. Several binaural noise reduction alg…