Birzeit Arabic Dialect Identification System for the 2018 VarDial Challenge
This paper describes our Automatic Dialect Recognition (ADI) system for the VarDial 2018 challenge, with the goal of distinguishing four major Arabic dialects, as well as Modern Standard Arabic (MSA). The training and development ADI VarDial 2018 data consists of 16,157 utterances, their words transcription, their phonetic transcriptions obtained with four non-Arabic phoneme recognizers and acoustic embedding data. Our overall system is a combination of four different systems. One system uses the words transcriptions and tries to recognize the speaker dialect by modeling the sequence of words for each dialect. Another system tries to recognize the dialect by modeling the phones sequence produced by non-Arabic phone recognizers, whereas, the other two systems use GMM trained on the acoustic features for recognizing the dialect. The best performance was achieved by the fused system which combines four systems together, with F1 micro of 68.77{\%}.
Code (0)
등록된 구현이 없습니다.
Tasks
Dialect IdentificationLanguage ModelingLanguage ModellingSpeech RecognitionSimilar Papers 제목 키워드 기반
Classifier Ensembles for Dialect and Language Variety Identification
In this paper we present ensemble-based systems for dialect and language variety identification using the datasets made available by the organizers of the VarDial Evaluation Campaign 2018. We present a system developed t…
Dialect IdentificationFindings of the VarDial Evaluation Campaign 2017
We present the results of the VarDial Evaluation Campaign on Natural Language Processing (NLP) for Similar Languages, Varieties and Dialects, which we organized as part of the fourth edition of the VarDial workshop at EA…
Dependency ParsingDialect IdentificationLanguage IdentificationLanguage Identification and Morphosyntactic Tagging: The Second VarDial Evaluation Campaign
We present the results and the findings of the Second VarDial Evaluation Campaign on Natural Language Processing (NLP) for Similar Languages, Varieties and Dialects. The campaign was organized as part of the fifth editio…
Dependency ParsingDialect IdentificationLanguage IdentificationArabic Dialect Identification Using iVectors and ASR Transcripts
This paper presents the systems submitted by the MAZA team to the Arabic Dialect Identification (ADI) shared task at the VarDial Evaluation Campaign 2017. The goal of the task is to evaluate computational models to ident…
Dialect IdentificationMachine TranslationThe GW/LT3 VarDial 2016 Shared Task System for Dialects and Similar Languages Detection
This paper describes the GW/LT3 contribution to the 2016 VarDial shared task on the identification of similar languages (task 1) and Arabic dialects (task 2). For both tasks, we experimented with Logistic Regression and …
Feature EngineeringregressionTask 2