ArabicNLU 2024: The First Arabic Natural Language Understanding Shared Task
This paper presents an overview of the Arabic Natural Language Understanding (ArabicNLU 2024) shared task, focusing on two subtasks: Word Sense Disambiguation (WSD) and Location Mention Disambiguation (LMD). The task aimed to evaluate the ability of automated systems to resolve word ambiguity and identify locations mentioned in Arabic text. We provided participants with novel datasets, including a sense-annotated corpus for WSD, called SALMA with approximately 34k annotated tokens, and the IDRISI-DA dataset with 3,893 annotations and 763 unique location mentions. These are challenging tasks. Out of the 38 registered teams, only three teams participated in the final evaluation phase, with the highest accuracy being 77.8% for WSD and the highest MRR@1 being 95.0% for LMD. The shared task not only facilitated the evaluation and comparison of different techniques, but also provided valuable insights and resources for the continued advancement of Arabic NLU technologies.
Code (0)
등록된 구현이 없습니다.
Tasks
Natural Language UnderstandingWord Sense DisambiguationSimilar Papers 제목 키워드 기반
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
Commonsense validation evaluates whether a sentence aligns with everyday human understanding, a critical capability for developing robust natural language understanding systems. While substantial progress has been made i…
Natural Language UnderstandingX-PuDu at SemEval-2022 Task 6: Multilingual Learning for English and Arabic Sarcasm Detection
Detecting sarcasm and verbal irony from people's subjective statements is crucial to understanding their intended meanings and real sentiments and positions in social scenarios. This paper describes the X-PuDu system tha…
Multi-Label ClassificationMUlTI-LABEL-ClASSIFICATIONNatural Language UnderstandingSarcasm Detection+1Is this sentence valid? An Arabic Dataset for Commonsense Validation
The commonsense understanding and validation remains a challenging task in the field of natural language understanding. Therefore, several research papers have been published that studied the capability of proposed syste…
Natural Language UnderstandingSentencevalidAraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization
Like most natural language understanding and generation tasks, state-of-the-art models for summarization are transformer-based sequence-to-sequence architectures that are pretrained on large corpora. While most existing …
Abstractive Text SummarizationDecoderNatural Language UnderstandingAraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization
Like most natural language understanding and generation tasks, state-of-the-art models for summarization are transformer-based sequence-to-sequence architectures that are pretrained on large corpora. While most existing …
Abstractive Text SummarizationDecoderNatural Language Understanding