paper-with-me

홈 › Papers

Building the Emirati Arabic FrameNet

2020-05-01 · LREC 2020 5 · Andrew Gargett, Tommi Leung

The Emirati Arabic FrameNet (EAFN) project aims to initiate a FrameNet for Emirati Arabic, utilizing the Emirati Arabic Corpus. The goal is to create a resource comparable to the initial stages of the Berkeley FrameNet. The project is divided into manual and automatic tracks, based on the predominant techniques being used to collect frames in each track. Work on the EAFN is progressing, and we here report on initial results for annotations and evaluation. The EAFN project aims to provide a general semantic resource for the Arabic language, sure to be of interest to researchers from general linguistics to natural language processing. As we report here, the EAFN is well on target for the first release of data in the coming year.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Mixat: A Data Set of Bilingual Emirati-English Speech

2024-05-04 · Maryam Al Ali, Hanan Aldarmaki

This paper introduces Mixat: a dataset of Emirati speech code-mixed with English. Mixat was developed to address the shortcomings of current speech recognition resources when applied to Emirati speech, and in particular,…

speech-recognitionSpeech Recognition

A Morphologically Annotated Corpus of Emirati Arabic

2018-05-01 · LREC 2018 5 · Salam Khalifa, Nizar Habash, Fadhl Eryani, Ossama Obeid 외
LemmatizationMachine TranslationMorphological AnalysisPart-Of-Speech Tagging

ArabicDialectHub: A Cross-Dialectal Arabic Learning Resource and Platform

2026-01-30 · Salem Lahlou arxiv

We present ArabicDialectHub, a cross-dialectal Arabic learning resource comprising 552 phrases across six varieties (Moroccan Darija, Lebanese, Syrian, Emirati, Saudi, and MSA) and an interactive web platform. Phrases we…

Distractor Generation

Ramsa: A Large Sociolinguistically Rich Emirati Arabic Speech Corpus for ASR and TTS

2026-03-09 · Rania Al-Sabbagh arxiv

Ramsa is a developing 41-hour speech corpus of Emirati Arabic designed to support sociolinguistic research and low-resource language technologies. It contains recordings from structured interviews with native speakers an…

Speech Recognition

DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models

2025-10-31 · Malik H. Altakrori, Nizar Habash, Abed Alhakim Freihat, Younes Samih 외 arxiv

We present DialectalArabicMMLU, a new benchmark for evaluating the performance of large language models (LLMs) across Arabic dialects. While recently developed Arabic and multilingual benchmarks have advanced LLM evaluat…