paper-with-me

Papers

STT4SG-350: A Speech Corpus for All Swiss German Dialect Regions

2023-05-30 · Michel Plüss, Jan Deriu, Yanick Schraner, Claudio Paonessa, Julia Hartmann, Larissa Schmidt, Christian Scheller, Manuela Hürlimann, Tanja Samardžić, Manfred Vogel, Mark Cieliebak

We present STT4SG-350 (Speech-to-Text for Swiss German), a corpus of Swiss German speech, annotated with Standard German text at the sentence level. The data is collected using a web app in which the speakers are shown Standard German sentences, which they translate to Swiss German and record. We make the corpus publicly available. It contains 343 hours of speech from all dialect regions and is the largest public speech corpus for Swiss German to date. Application areas include automatic speech recognition (ASR), text-to-speech, dialect identification, and speaker recognition. Dialect information, age group, and gender of the 316 speakers are provided. Genders are equally represented and the corpus includes speakers of all ages. Roughly the same amount of speech is provided per dialect region, which makes the corpus ideally suited for experiments with speech technology for different dialects. We provide training, validation, and test splits of the data. The test set consists of the same spoken sentences for each dialect region and allows a fair evaluation of the quality of speech technologies in different dialects. We train an ASR model on the training set and achieve an average BLEU score of 74.7 on the test set. The model beats the best published BLEU scores on 2 other Swiss German ASR test sets, demonstrating the quality of the corpus.

📄 PDF Abstract BibTeX arXiv:2305.18855

Code (0)

등록된 구현이 없습니다.

Tasks

AllAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect IdentificationSentenceSpeaker Recognitionspeech-recognitionSpeech RecognitionSpeech-to-Texttext-to-speechText to Speech

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

SwissGPC v1.0 -- The Swiss German Podcasts Corpus

2025-09-24 · Samuel Stucki, Mark Cieliebak, Jan Deriu arxiv

We present SwissGPC v1.0, the first mid-to-large-scale corpus of spontaneous Swiss German speech, developed to support research in ASR, TTS, dialect identification, and related fields. The dataset consists of links to ta…

SDS-200: A Swiss German Speech to Standard German Text Corpus

2022-05-19 · LREC 2022 6 · Michel Plüss, Manuela Hürlimann, Marc Cuny, Alla Stöckli 외

We present SDS-200, a corpus of Swiss German dialectal speech with Standard German text translations, annotated with dialect, age, and gender information of the speakers. The dataset allows for training speech translatio…

Speech SynthesisTranslation

SwissDial: Parallel Multidialectal Corpus of Spoken Swiss German

2021-03-21 · Pelin Dogan-Schönberger, Julian Mäder, Thomas Hofmann

Swiss German is a dialect continuum whose natively acquired dialects significantly differ from the formal variety of the language. These dialects are mostly used for verbal communication and do not have standard orthogra…

Speech Synthesis

Dialect Transfer for Swiss German Speech Translation

2023-10-13 · Claudio Paonessa, Yanick Schraner, Jan Deriu, Manuela Hürlimann 외

This paper investigates the challenges in building Swiss German speech translation systems, specifically focusing on the impact of dialect diversity and differences between Swiss German and Standard German. Swiss German …

DiversityTranslation

ArchiMob - A Corpus of Spoken Swiss German

2016-05-01 · LREC 2016 5 · Tanja Samard{\v{z}}i{\'c}, Yves Scherrer, Elvira Glaser

Swiss dialects of German are, unlike most dialects of well standardised languages, widely used in everyday communication. Despite this fact, automatic processing of Swiss German is still a considerable challenge due to t…

Machine TranslationPart-Of-Speech TaggingTranslation