German Dialect Identification and Mapping for Preservation and Recovery
Many linguistic projects which focus on dialects do collection of audio data, analysis, and linguistic interpretation on the data. The outcomes of such projects are good language resources because dialects are among less-resources languages as most of them are oral traditions. Our project Dialektatlas Mittleres Westdeutschland (DMW) 1 focuses on the study of German language varieties through collection of audio data of words and phrases which are selected by linguistic experts based on the linguistic significance of the words (and phrases) to distinguish dialects among each other. We used a total of 7,814 audio snippets of the words and phrases of eight different dialects from middle west Germany. We employed a multilabel classification approach to address the problem of dialect mapping using Support Vector Machine (SVM) algorithm. The experimental result showed a promising accuracy of 87%.
Code (0)
등록된 구현이 없습니다.
Tasks
Dialect IdentificationSimilar Papers 제목 키워드 기반
TwistBytes - Identification of Cuneiform Languages and German Dialects at VarDial 2019
We describe our approaches for the German Dialect Identification (GDI) and the Cuneiform Language Identification (CLI) tasks at the VarDial Evaluation Campaign 2019. The goal was to identify dialects of Swiss German in G…
Dialect IdentificationLanguage IdentificationGerman Dialect Identification in Interview Transcriptions
This paper presents three systems submitted to the German Dialect Identification (GDI) task at the VarDial Evaluation Campaign 2017. The task consists of training models to identify the dialect of Swiss-German speech tra…
Dialect IdentificationMachine TranslationGerman Dialect Identification Using Classifier Ensembles
In this paper we present the GDI_classification entry to the second German Dialect Identification (GDI) shared task organized within the scope of the VarDial Evaluation Campaign 2018. We present a system based on SVM cla…
Dialect IdentificationTwist Bytes - German Dialect Identification with Data Mining Optimization
We describe our approaches used in the German Dialect Identification (GDI) task at the VarDial Evaluation Campaign 2018. The goal was to identify to which out of four dialects spoken in German speaking part of Switzerlan…
Dialect IdentificationSentenceSTT4SG-350: A Speech Corpus for All Swiss German Dialect Regions
We present STT4SG-350 (Speech-to-Text for Swiss German), a corpus of Swiss German speech, annotated with Standard German text at the sentence level. The data is collected using a web app in which the speakers are shown S…
AllAutomatic Speech RecognitionAutomatic Speech Recognition (ASR)Dialect Identification+7