RescueSpeech: A German Corpus for Speech Recognition in Search and Rescue Domain
Despite the recent advancements in speech recognition, there are still difficulties in accurately transcribing conversational and emotional speech in noisy and reverberant acoustic environments. This poses a particular challenge in the search and rescue (SAR) domain, where transcribing conversations among rescue team members is crucial to support real-time decision-making. The scarcity of speech data and associated background noise in SAR scenarios make it difficult to deploy robust speech recognition systems. To address this issue, we have created and made publicly available a German speech dataset called RescueSpeech. This dataset includes real speech recordings from simulated rescue exercises. Additionally, we have released competitive training recipes and pre-trained models. Our study highlights that the performance attained by state-of-the-art methods in this challenging scenario is still far from reaching an acceptable level.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingRobust Speech Recognitionspeech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
LibriVoxDeEn: A Corpus for German-to-English Speech Translation and German Speech Recognition
We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audiobooks. The speech translation data consist of 110 hours of audio material aligned to over 50k pa…
Sentencespeech-recognitionSpeech RecognitionTranslationOpen Source German Distant Speech Recognition: Corpus and Acoustic Model
We present a new freely available corpus for German distant speech recognition and report speaker-independent word error rate (WER) results for two open source speech recognizers trained on this corpus. The corpus has be…
Distant Speech Recognitionspeech-recognitionSpeech RecognitionSTT4SG-350: A Speech Corpus for All Swiss German Dialect Regions
We present STT4SG-350 (Speech-to-Text for Swiss German), a corpus of Swiss German speech, annotated with Standard German text at the sentence level. The data is collected using a web app in which the speakers are shown S…
AllAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect Identification+7Exploiting the large-scale German Broadcast Corpus to boost the Fraunhofer IAIS Speech Recognition System
In this paper we describe the large-scale German broadcast corpus (GER-TV1000h) containing more than 1,000 hours of transcribed speech data. This corpus is unique in the German language corpora domain and enables signifi…
Acoustic ModellingAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Language Modelling+5RWTH-PHOENIX-Weather: A Large Vocabulary Sign Language Recognition and Translation Corpus
This paper introduces the RWTH-PHOENIX-Weather corpus, a video-based, large vocabulary corpus of German Sign Language suitable for statistical sign language recognition and translation. In contrastto most available sign …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Machine TranslationSentence+4