paper-with-me

홈 › Papers

AL-QASIDA: Analyzing LLM Quality and Accuracy Systematically in Dialectal Arabic

2024-12-05 · Nathaniel R. Robinson, Shahd Abdelmoneim, Kelly Marchisio, Sebastian Ruder

Dialectal Arabic (DA) varieties are under-served by language technologies, particularly large language models (LLMs). This trend threatens to exacerbate existing social inequalities and limits LLM applications, yet the research community lacks operationalized performance measurements in DA. We present a framework that comprehensively assesses LLMs' DA modeling capabilities across four dimensions: fidelity, understanding, quality, and diglossia. We evaluate nine LLMs in eight DA varieties and provide practical recommendations. Our evaluation suggests that LLMs do not produce DA as well as they understand it, not because their DA fluency is poor, but because they are reluctant to generate DA. Further analysis suggests that current post-training can contribute to bias against DA, that few-shot examples can overcome this deficiency, and that otherwise no measurable features of input text correlate well with LLM DA performance.

📄 PDF Abstract BibTeX arXiv:2412.04193

Code (0)

등록된 구현이 없습니다.

Tasks

Language Modelling

Similar Papers 제목 키워드 기반

Ara-HOPE: Human-Centric Post-Editing Evaluation for Dialectal Arabic to Modern Standard Arabic Translation

2025-12-25 · Abdullah Alabdullah, Lifeng Han, Chenghua Lin arxiv

Dialectal Arabic to Modern Standard Arabic (DA-MSA) translation is a challenging task in Machine Translation (MT) due to significant lexical, syntactic, and semantic divergences between Arabic dialects and MSA. Existing …

Machine Translation

EgyBERT: A Large Language Model Pretrained on Egyptian Dialect Corpora

2024-08-07 · Faisal Qarah

This study presents EgyBERT, an Arabic language model pretrained on 10.4 GB of Egyptian dialectal texts. We evaluated EgyBERT's performance by comparing it with five other multidialect Arabic language models across 10 ev…

Language ModelingLanguage ModellingLarge Language Model

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

2026-04-07 · Jason Lucas, Matt Murtagh, Ali Al-Lawati, Uchendu Uchendu 외 arxiv

Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), leaving their robustness to dialectal variation unexplored. We present DIA…

A Catalog of Basque Dialectal Resources: Online Collections and Standard-to-Dialectal Adaptations

2026-03-26 · Jaione Bengoetxea, Itziar Gonzalez-Dios, Rodrigo Agerri arxiv

Recent research on dialectal NLP has identified data scarcity as a primary limitation. To address this limitation, this paper presents a catalog of contemporary Basque dialectal data and resources, offering a systematic …

Natural Language Inference

Bailing-TTS: Chinese Dialectal Speech Synthesis Towards Human-like Spontaneous Representation

2024-08-01 · Xinhan Di, Zihao Chen, Yunming Liang, Junjie Zheng 외

Large-scale text-to-speech (TTS) models have made significant progress recently.However, they still fall short in the generation of Chinese dialectal speech. Toaddress this, we propose Bailing-TTS, a family of large-scal…

Representation LearningSpeech Synthesistext-to-speechText to Speech