paper-with-me

홈 › Papers

Deep Learning Based Speech Beamforming

2018-02-15 · Kaizhi Qian, Yang Zhang, Shiyu Chang, Xuesong Yang, Dinei Florencio, Mark Hasegawa-Johnson

Multi-channel speech enhancement with ad-hoc sensors has been a challenging task. Speech model guided beamforming algorithms are able to recover natural sounding speech, but the speech models tend to be oversimplified or the inference would otherwise be too complicated. On the other hand, deep learning based enhancement approaches are able to learn complicated speech distributions and perform efficient inference, but they are unable to deal with variable number of input channels. Also, deep learning approaches introduce a lot of errors, particularly in the presence of unseen noise types and settings. We have therefore proposed an enhancement framework called DEEPBEAM, which combines the two complementary classes of algorithms. DEEPBEAM introduces a beamforming filter to produce natural sounding speech, but the filter coefficients are determined with the help of a monaural speech enhancement neural network. Experiments on synthetic and real-world data show that DEEPBEAM is able to produce clean, dry and natural sounding speech, and is robust against unseen noise.

📄 PDF Abstract BibTeX arXiv:1802.05383

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningSpeech Enhancement

Similar Papers 제목 키워드 기반

Lightweight Speech Enhancement in Unseen Noisy and Reverberant Conditions using KISS-GEV Beamforming

2021-10-06 · Thomas Bernard, François Grondin

This paper introduces a new method referred to as KISS-GEV (for Keep It Super Simple Generalized eigenvalue) beamforming. While GEV beamforming usually relies on deep neural network for estimating target and noise time-f…

Speech Enhancement

A unified multichannel far-field speech recognition system: combining neural beamforming with attention based end-to-end model

2024-01-05 · Dongdi Zhao, Jianbo Ma, Lu Lu, Jinke Li 외

Far-field speech recognition is a challenging task that conventionally uses signal processing beamforming to attack noise and interference problem. But the performance has been found usually limited due to heavy reliance…

Speech Enhancementspeech-recognitionSpeech Recognition

Deep Long Short-Term Memory Adaptive Beamforming Networks For Multichannel Robust Speech Recognition

2017-11-21 · Zhong Meng, Shinji Watanabe, John R. Hershey, Hakan Erdogan

Far-field speech recognition in noisy and reverberant conditions remains a challenging problem despite recent deep learning breakthroughs. This problem is commonly addressed by acquiring a speech signal from multiple mic…

Robust Speech Recognitionspeech-recognitionSpeech Recognition

Speaker Adapted Beamforming for Multi-Channel Automatic Speech Recognition

2018-06-19 · Tobias Menne, Ralf Schlüter, Hermann Ney

This paper presents, in the context of multi-channel ASR, a method to adapt a mask based, statistically optimal beamforming approach to a speaker of interest. The beamforming vector of the statistically optimal beamforme…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Run-Time Adaptation of Neural Beamforming for Robust Speech Dereverberation and Denoising

2024-10-30 · Yoto Fujita, Aditya Arie Nugraha, Diego Di Carlo, Yoshiaki Bando 외

This paper describes speech enhancement for realtime automatic speech recognition (ASR) in real environments. A standard approach to this task is to use neural beamforming that can work efficiently in an online manner. I…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingSpeech Dereverberation+3