paper-with-me

Papers

Arabic Dialect Identification Using BERT-Based Domain Adaptation

2020-11-13 · COLING (WANLP) 2020 12 · Ahmad Beltagy, Abdelrahman Wael, Omar ElSherief

Arabic is one of the most important and growing languages in the world. With the rise of social media platforms such as Twitter, Arabic spoken dialects have become more in use. In this paper, we describe our approach on the NADI Shared Task 1 that requires us to build a system to differentiate between different 21 Arabic dialects, we introduce a deep learning semi-supervised fashion approach along with pre-processing that was reported on NADI shared Task 1 Corpus. Our system ranks 4th in NADI's shared task competition achieving a 23.09% F1 macro average score with a simple yet efficient approach to differentiating between 21 Arabic Dialects given tweets.

📄 PDF Abstract BibTeX arXiv:2011.06977

Code (0)

등록된 구현이 없습니다.

Tasks

Dialect IdentificationDomain Adaptation

Similar Papers 제목 키워드 기반

Multi-Dialect Arabic BERT for Country-Level Dialect Identification

2020-07-10 · COLING (WANLP) 2020 12 · Bashar Talafha, Mohammad Ali, Muhy Eddin Za'ter, Haitham Seelawi 외

Arabic dialect identification is a complex problem for a number of inherent properties of the language itself. In this paper, we present the experiments conducted, and the models developed by our competing team, Mawdoo3 …

Dialect IdentificationLanguage ModelingLanguage Modelling

Domain Adaptation for Arabic Cross-Domain and Cross-Dialect Sentiment Analysis from Contextualized Word Embedding

2021-06-01 · NAACL 2021 4 · Abdellah El Mekki, Abdelkader El Mahdaouy, Ismail Berrada, Ahmed Khoumsi

Finetuning deep pre-trained language models has shown state-of-the-art performances on a wide range of Natural Language Processing (NLP) applications. Nevertheless, their generalization performance drops under domain shi…

Domain AdaptationSentiment AnalysisTransfer LearningUnsupervised Domain Adaptation

Arabic dialect identification: An Arabic-BERT model with data augmentation and ensembling strategy

2020-12-01 · COLING (WANLP) 2020 12 · Kamel Gaanoun, Imade Benelallam

This paper presents the ArabicProcessors team’s deep learning system designed for the NADI 2020 Subtask 1 (country-level dialect identification) and Subtask 2 (province-level dialect identification). We used Arabic-Bert …

Data AugmentationDialect Identification

BERT-based Multi-Task Model for Country and Province Level Modern Standard Arabic and Dialectal Arabic Identification

2021-06-23 · Abdellah El Mekki, Abdelkader El Mahdaouy, Kabil Essefar, Nabil El Mamoun 외

Dialect and standard language identification are crucial tasks for many Arabic natural language processing applications. In this paper, we present our deep learning-based system, submitted to the second NADI shared task …

Language IdentificationMulti-Task Learning

Adapting MARBERT for Improved Arabic Dialect Identification: Submission to the NADI 2021 Shared Task

2021-03-01 · EACL (WANLP) 2021 4 · Badr AlKhamissi, Mohamed Gabr, Muhammad ElNokrashy, Khaled Essam

In this paper, we tackle the Nuanced Arabic Dialect Identification (NADI) shared task (Abdul-Mageed et al., 2021) and demonstrate state-of-the-art results on all of its four subtasks. Tasks are to identify the geographic…

Dialect Identification