paper-with-me

홈 › Papers

A network of deep neural networks for distant speech recognition

2017-03-23 · Mirco Ravanelli, Philemon Brakel, Maurizio Omologo, Yoshua Bengio

Despite the remarkable progress recently made in distant speech recognition, state-of-the-art technology still suffers from a lack of robustness, especially when adverse acoustic conditions characterized by non-stationary noises and reverberation are met. A prominent limitation of current systems lies in the lack of matching and communication between the various technologies involved in the distant speech recognition process. The speech enhancement and speech recognition modules are, for instance, often trained independently. Moreover, the speech enhancement normally helps the speech recognizer, but the output of the latter is not commonly used, in turn, to improve the speech enhancement. To address both concerns, we propose a novel architecture based on a network of deep neural networks, where all the components are jointly trained and better cooperate with each other thanks to a full communication scheme between them. Experiments, conducted using different datasets, tasks and acoustic conditions, revealed that the proposed framework can overtake other competitive solutions, including recent joint training approaches.

📄 PDF Abstract BibTeX arXiv:1703.08002

Code (0)

등록된 구현이 없습니다.

Tasks

Distant Speech RecognitionSpeech Enhancementspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

A Study of Enhancement, Augmentation, and Autoencoder Methods for Domain Adaptation in Distant Speech Recognition

2018-06-13 · Hao Tang, Wei-Ning Hsu, Francois Grondin, James Glass

Speech recognizers trained on close-talking speech do not generalize to distant speech and the word error rate degradation can be as large as 40% absolute. Most studies focus on tackling distant speech recognition as a s…

Data AugmentationDistant Speech RecognitionDomain AdaptationSpeech Enhancement+2

End-to-end attention-based distant speech recognition with Highway LSTM

2016-10-17 · Hassan Taherian

End-to-end attention-based models have been shown to be competitive alternatives to conventional DNN-HMM models in the Speech Recognition Systems. In this paper, we extend existing end-to-end attention-based models that …

Distant Speech Recognitionspeech-recognitionSpeech Recognition

Open Source German Distant Speech Recognition: Corpus and Acoustic Model

2015-12-11 · International Conference on Text, Speech, and Dialogue 2015 12 · Stephan Radeck-Arneth, Benjamin Milde, Arvid Lange, Evandro Gouvea 외

We present a new freely available corpus for German distant speech recognition and report speaker-independent word error rate (WER) results for two open source speech recognizers trained on this corpus. The corpus has be…

Distant Speech Recognitionspeech-recognitionSpeech Recognition

Analyzing Large Receptive Field Convolutional Networks for Distant Speech Recognition

2019-10-15 · Salar Jafarlou, Soheil Khorram, Vinay Kothapally, John H. L. Hansen

Despite significant efforts over the last few years to build a robust automatic speech recognition (ASR) system for different acoustic settings, the performance of the current state-of-the-art technologies significantly …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Distant Speech Recognitionspeech-recognition+1

BridgeNets: Student-Teacher Transfer Learning Based on Recursive Neural Networks and its Application to Distant Speech Recognition

2017-10-27 · Jaeyoung Kim, Mostafa El-Khamy, Jungwon Lee

Despite the remarkable progress achieved on automatic speech recognition, recognizing far-field speeches mixed with various noise sources is still a challenging task. In this paper, we introduce novel student-teacher tra…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingDistant Speech Recognition+3