paper-with-me

홈 › Papers

A Survey of Code-switched Arabic NLP: Progress, Challenges, and Future Directions

2025-01-23 · Injy Hamed, Caroline Sabty, Slim Abdennadher, Ngoc Thang Vu, Thamar Solorio, Nizar Habash

Language in the Arab world presents a complex diglossic and multilingual setting, involving the use of Modern Standard Arabic, various dialects and sub-dialects, as well as multiple European languages. This diverse linguistic landscape has given rise to code-switching, both within Arabic varieties and between Arabic and foreign languages. The widespread occurrence of code-switching across the region makes it vital to address these linguistic needs when developing language technologies. In this paper, we provide a review of the current literature in the field of code-switched Arabic NLP, offering a broad perspective on ongoing efforts, challenges, research gaps, and recommendations for future research directions.

📄 PDF Abstract BibTeX arXiv:2501.13419

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Overview for the Second Shared Task on Language Identification in Code-Switched Data

2019-09-28 · WS 2016 11 · Giovanni Molina, Fahad AlGhamdi, Mahmoud Ghoneim, Abdelati Hawwari 외

We present an overview of the second shared task on language identification in code-switched data. For the shared task, we had code-switched data from two different language pairs: Modern Standard Arabic-Dialectal Arabic…

Language IdentificationSingle Particle Analysis

ArzEn-LLM: Code-Switched Egyptian Arabic-English Translation and Speech Recognition Using LLMs

2024-06-26 · Ahmed Heakl, Youssef Zaghloul, Mennatullah Ali, Rania Hossam 외

Motivated by the widespread increase in the phenomenon of code-switching between Egyptian Arabic and English in recent times, this paper explores the intricacies of machine translation (MT) and automatic speech recogniti…

ArzEn Code-switched Translation to araArzEn Code-switched Translation to engArzEn Speech RecognitionAutomatic Speech Recognition+9

Named Entity Recognition on Code-Switched Data: Overview of the CALCS 2018 Shared Task

2019-06-10 · WS 2018 7 · Gustavo Aguilar, Fahad AlGhamdi, Victor Soto, Mona Diab 외

In the third shared task of the Computational Approaches to Linguistic Code-Switching (CALCS) workshop, we focus on Named Entity Recognition (NER) on code-switched social-media data. We divide the shared task into two co…

Diversitynamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Computational Approaches to Arabic-English Code-Switching

2024-10-17 · Caroline Sabty

Natural Language Processing (NLP) is a vital computational method for addressing language processing, analysis, and generation. NLP tasks form the core of many daily applications, from automatic text correction to speech…

Data AugmentationLanguage Identificationnamed-entity-recognitionNamed Entity Recognition+4

Contextual Embeddings for Arabic-English Code-Switched Data

2020-12-01 · COLING (WANLP) 2020 12 · Caroline Sabty, Mohamed Islam, Slim Abdennadher

Globalization has caused the rise of the code-switching phenomenon among multilingual societies. In Arab countries, code-switching between Arabic and English has become frequent, especially through social media platforms…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Question Answering+1