paper-with-me

Papers

Deep Learning Based Dereverberation of Temporal Envelopesfor Robust Speech Recognition

2020-08-07 · Anurenjan Purushothaman, Anirudh Sreeram, Rohit Kumar, Sriram Ganapathy

Automatic speech recognition in reverberant conditions is a challenging task as the long-term envelopes of the reverberant speech are temporally smeared. In this paper, we propose a neural model for enhancement of sub-band temporal envelopes for dereverberation of speech. The temporal envelopes are derived using the autoregressive modeling framework of frequency domain linear prediction (FDLP). The neural enhancement model proposed in this paper performs an envelop gain based enhancement of temporal envelopes and it consists of a series of convolutional and recurrent neural network layers. The enhanced sub-band envelopes are used to generate features for automatic speech recognition (ASR). The ASR experiments are performed on the REVERB challenge dataset as well as the CHiME-3 dataset. In these experiments, the proposed neural enhancement approach provides significant improvements over a baseline ASR system with beamformed audio (average relative improvements of 21% on the development set and about 11% on the evaluation set in word error rates for REVERB challenge dataset).

📄 PDF Abstract BibTeX arXiv:2008.03339

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Deep LearningRobust Speech Recognitionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Dereverberation of Autoregressive Envelopes for Far-field Speech Recognition

2021-08-12 · Anurenjan Purushothaman, Anirudh Sreeram, Rohit Kumar, Sriram Ganapathy

The task of speech recognition in far-field environments is adversely affected by the reverberant artifacts that elicit as the temporal smearing of the sub-band envelopes. In this paper, we develop a neural model for spe…

Speech Dereverberationspeech-recognitionSpeech Recognition

End-to-End Speech Recognition With Joint Dereverberation Of Sub-Band Autoregressive Envelopes

2021-08-09 · Rohit Kumar, Anurenjan Purushothaman, Anirudh Sreeram, Sriram Ganapathy

The end-to-end (E2E) automatic speech recognition (ASR) systems are often required to operate in reverberant conditions, where the long-term sub-band envelopes of the speech are temporally smeared. In this paper, we deve…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

TeCANet: Temporal-Contextual Attention Network for Environment-Aware Speech Dereverberation

2021-03-31 · Helin Wang, Bo Wu, LianWu Chen, Meng Yu 외

In this paper, we exploit the effective way to leverage contextual information to improve the speech dereverberation performance in real-world reverberant environments. We propose a temporal-contextual attention approach…

Room Impulse Response (RIR)Speech Dereverberation

Audio-visual multi-channel speech separation, dereverberation and recognition

2022-04-05 · Guinan Li, Jianwei Yu, Jiajun Deng, Xunying Liu 외

Despite the rapid advance of automatic speech recognition (ASR) technologies, accurate recognition of cocktail party speech characterised by the interference from overlapping speakers, background noise and room reverbera…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+2

Investigating Generative Adversarial Networks based Speech Dereverberation for Robust Speech Recognition

2018-03-27 · Ke Wang, Junbo Zhang, Sining Sun, Yujun Wang 외

We investigate the use of generative adversarial networks (GANs) in speech dereverberation for robust speech recognition. GANs have been recently studied for speech enhancement to remove additive noises, but there still …

Robust Speech RecognitionSpeech DereverberationSpeech Enhancementspeech-recognition+1