A Morphologically Annotated Corpus of Emirati Arabic
Code (0)
등록된 구현이 없습니다.
Tasks
LemmatizationMachine TranslationMorphological AnalysisPart-Of-Speech TaggingSimilar Papers 제목 키워드 기반
Building the Emirati Arabic FrameNet
The Emirati Arabic FrameNet (EAFN) project aims to initiate a FrameNet for Emirati Arabic, utilizing the Emirati Arabic Corpus. The goal is to create a resource comparable to the initial stages of the Berkeley FrameNet. …
Curras + Baladi: Towards a Levantine Corpus
The processing of the Arabic language is a complex field of research. This is due to many factors, including the complex and rich morphology of Arabic, its high degree of ambiguity, and the presence of several regional v…
Ramsa: A Large Sociolinguistically Rich Emirati Arabic Speech Corpus for ASR and TTS
Ramsa is a developing 41-hour speech corpus of Emirati Arabic designed to support sociolinguistic research and low-resource language technologies. It contains recordings from structured interviews with native speakers an…
Speech RecognitionMixat: A Data Set of Bilingual Emirati-English Speech
This paper introduces Mixat: a dataset of Emirati speech code-mixed with English. Mixat was developed to address the shortcomings of current speech recognition resources when applied to Emirati speech, and in particular,…
speech-recognitionSpeech RecognitionMorphologically Annotated Corpora and Morphological Analyzers for Moroccan and Sanaani Yemeni Arabic
We present new language resources for Moroccan and Sanaani Yemeni Arabic. The resources include corpora for each dialect which have been morphologically annotated, and morphological analyzers for each dialect which are d…