paper-with-me

Papers

Open Source German Distant Speech Recognition: Corpus and Acoustic Model

2015-12-11 · International Conference on Text, Speech, and Dialogue 2015 12 · Stephan Radeck-Arneth, Benjamin Milde, Arvid Lange, Evandro Gouvea, Stefan Radomski, Max Mühlhäuser, and Chris Biemann

We present a new freely available corpus for German distant speech recognition and report speaker-independent word error rate (WER) results for two open source speech recognizers trained on this corpus. The corpus has been recorded in a controlled environment with three different microphones at a distance of one meter. It comprises 180 different speakers with a total of 36 hours of audio recordings. We show recognition results with the open source toolkit Kaldi (20.5% WER) and PocketSphinx (39.6% WER) and make a complete open source solution for German distant speech recognition possible.

📄 PDF Abstract BibTeX

Code (1)

tudarmstadt-lt/kaldi-tuda-de 공식 구현

Tasks

Distant Speech Recognitionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Open Source Automatic Speech Recognition for German

2018-07-26 · Benjamin Milde, Arne Köhn

High quality Automatic Speech Recognition (ASR) is a prerequisite for speech-based applications and research. While state-of-the-art ASR software is freely available, the language dependent acoustic models are lacking fo…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

LibriVoxDeEn: A Corpus for German-to-English Speech Translation and German Speech Recognition

2019-10-17 · LREC 2020 5 · Benjamin Beilharz, Xin Sun, Sariya Karimova, Stefan Riezler

We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audiobooks. The speech translation data consist of 110 hours of audio material aligned to over 50k pa…

Sentencespeech-recognitionSpeech RecognitionTranslation

IMS-Speech: A Speech to Text Tool

2019-08-13 · Pavel Denisov, Ngoc Thang Vu

We present the IMS-Speech, a web based tool for German and English speech transcription aiming to facilitate research in various disciplines which require accesses to lexical information in spoken language materials. Thi…

speech-recognitionSpeech RecognitionSpeech-to-Text

Reconnaissance automatique de la parole distante dans un habitat intelligent : m\'ethodes multi-sources en conditions r\'ealistes (Distant Speech Recognition in a Smart Home : Comparison of Several Multisource ASRs in Realistic Conditions) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Benjamin Lecouteux, Michel Vacher, Fran{\c{c}}ois Portet
Distant Speech Recognitionspeech-recognitionSpeech Recognition

CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings

2020-04-20 · Shinji Watanabe, Michael Mandel, Jon Barker, Emmanuel Vincent 외

Following the success of the 1st, 2nd, 3rd, 4th and 5th CHiME challenges we organize the 6th CHiME Speech Separation and Recognition Challenge (CHiME-6). The new challenge revisits the previous CHiME-5 challenge and furt…

speaker-diarizationSpeaker DiarizationSpeech Enhancementspeech-recognition+2