paper-with-me

Papers

Two-Staged Acoustic Modeling Adaption for Robust Speech Recognition by the Example of German Oral History Interviews

2019-08-19 · Michael Gref, Christoph Schmidt, Sven Behnke, Joachim köhler

In automatic speech recognition, often little training data is available for specific challenging tasks, but training of state-of-the-art automatic speech recognition systems requires large amounts of annotated speech. To address this issue, we propose a two-staged approach to acoustic modeling that combines noise and reverberation data augmentation with transfer learning to robustly address challenges such as difficult acoustic recording conditions, spontaneous speech, and speech of elderly people. We evaluate our approach using the example of German oral history interviews, where a relative average reduction of the word error rate by 19.3% is achieved.

📄 PDF Abstract BibTeX arXiv:1908.06709

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationRobust Speech Recognitionspeech-recognitionSpeech RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

Multi-Staged Cross-Lingual Acoustic Model Adaption for Robust Speech Recognition in Real-World Applications - A Case Study on German Oral History Interviews

2020-05-01 · LREC 2020 5 · Michael Gref, Oliver Walter, Christoph Schmidt, Sven Behnke 외

While recent automatic speech recognition systems achieve remarkable performance when large amounts of adequate, high quality annotated speech data is used for training, the same systems often only achieve an unsatisfact…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Robust Speech Recognitionspeech-recognition+1

Improving Speech Recognition for the Elderly: A New Corpus of Elderly Japanese Speech and Investigation of Acoustic Modeling for Speech Recognition

2020-05-01 · LREC 2020 5 · Meiko Fukuda, Hiromitsu Nishizaki, Yurie Iribe, Ryota Nishimura 외

In an aging society like Japan, a highly accurate speech recognition system is needed for use in electronic devices for the elderly, but this level of accuracy cannot be obtained using conventional speech recognition sys…

speech-recognitionSpeech Recognition

Bridging the Gap Between Monaural Speech Enhancement and Recognition with Distortion-Independent Acoustic Modeling

2019-03-11 · Peidong Wang, Ke Tan, DeLiang Wang

Monaural speech enhancement has made dramatic advances since the introduction of deep learning a few years ago. Although enhanced speech has been demonstrated to have better intelligibility and quality for human listener…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+1

Large-Scale Mixed-Bandwidth Deep Neural Network Acoustic Modeling for Automatic Speech Recognition

2019-07-10 · Khoi-Nguyen C. Mac, Xiaodong Cui, Wei zhang, Michael Picheny

In automatic speech recognition (ASR), wideband (WB) and narrowband (NB) speech signals with different sampling rates typically use separate acoustic models. Therefore mixed-bandwidth (MB) acoustic modeling has important…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Bandwidth Extensionspeech-recognition+1

Automatic Speech Recognition on a Firefighter TETRA Broadcast Channel

2012-05-01 · LREC 2012 5 · Daniel Stein, Bela Usabaev

For a reliable keyword extraction on firefighter radio communication, a strong automatic speech recognition system is needed. However, real-life data poses several challenges like a distorted voice signal, background noi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Keyword ExtractionLanguage Modelling+3