paper-with-me

홈 › Papers

RescueSpeech: A German Corpus for Speech Recognition in Search and Rescue Domain

2023-06-06 · Sangeet Sagar, Mirco Ravanelli, Bernd Kiefer, Ivana Kruijff Korbayova, Josef van Genabith

Despite the recent advancements in speech recognition, there are still difficulties in accurately transcribing conversational and emotional speech in noisy and reverberant acoustic environments. This poses a particular challenge in the search and rescue (SAR) domain, where transcribing conversations among rescue team members is crucial to support real-time decision-making. The scarcity of speech data and associated background noise in SAR scenarios make it difficult to deploy robust speech recognition systems. To address this issue, we have created and made publicly available a German speech dataset called RescueSpeech. This dataset includes real speech recordings from simulated rescue exercises. Additionally, we have released competitive training recipes and pre-trained models. Our study highlights that the performance attained by state-of-the-art methods in this challenging scenario is still far from reaching an acceptable level.

📄 PDF Abstract BibTeX arXiv:2306.04054

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingRobust Speech Recognitionspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

LibriVoxDeEn: A Corpus for German-to-English Speech Translation and German Speech Recognition

2019-10-17 · LREC 2020 5 · Benjamin Beilharz, Xin Sun, Sariya Karimova, Stefan Riezler

We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audiobooks. The speech translation data consist of 110 hours of audio material aligned to over 50k pa…

Sentencespeech-recognitionSpeech RecognitionTranslation

Open Source German Distant Speech Recognition: Corpus and Acoustic Model

2015-12-11 · International Conference on Text, Speech, and Dialogue 2015 12 · Stephan Radeck-Arneth, Benjamin Milde, Arvid Lange, Evandro Gouvea 외

We present a new freely available corpus for German distant speech recognition and report speaker-independent word error rate (WER) results for two open source speech recognizers trained on this corpus. The corpus has be…

Distant Speech Recognitionspeech-recognitionSpeech Recognition

STT4SG-350: A Speech Corpus for All Swiss German Dialect Regions

2023-05-30 · Michel Plüss, Jan Deriu, Yanick Schraner, Claudio Paonessa 외

We present STT4SG-350 (Speech-to-Text for Swiss German), a corpus of Swiss German speech, annotated with Standard German text at the sentence level. The data is collected using a web app in which the speakers are shown S…

AllAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect Identification+7

Exploiting the large-scale German Broadcast Corpus to boost the Fraunhofer IAIS Speech Recognition System

2014-05-01 · LREC 2014 5 · Michael Stadtschnitzer, Jochen Schwenninger, Daniel Stein, Joachim Koehler

In this paper we describe the large-scale German broadcast corpus (GER-TV1000h) containing more than 1,000 hours of transcribed speech data. This corpus is unique in the German language corpora domain and enables signifi…

Acoustic ModellingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Modelling+5

RWTH-PHOENIX-Weather: A Large Vocabulary Sign Language Recognition and Translation Corpus

2012-05-01 · LREC 2012 5 · Jens Forster, Christoph Schmidt, Thomas Hoyoux, Oscar Koller 외

This paper introduces the RWTH-PHOENIX-Weather corpus, a video-based, large vocabulary corpus of German Sign Language suitable for statistical sign language recognition and translation. In contrastto most available sign …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSentence+4