paper-with-me

홈 › Papers

Multi-VALUE: A Framework for Cross-Dialectal English NLP

2022-12-15 · Caleb Ziems, William Held, Jingfeng Yang, Jwala Dhamala, Rahul Gupta, Diyi Yang

Dialect differences caused by regional, social, and economic factors cause performance discrepancies for many groups of language technology users. Inclusive and equitable language technology must critically be dialect invariant, meaning that performance remains constant over dialectal shifts. Current systems often fall short of this ideal since they are designed and tested on a single dialect: Standard American English (SAE). We introduce a suite of resources for evaluating and achieving English dialect invariance. The resource is called Multi-VALUE, a controllable rule-based translation system spanning 50 English dialects and 189 unique linguistic features. Multi-VALUE maps SAE to synthetic forms of each dialect. First, we use this system to stress tests question answering, machine translation, and semantic parsing. Stress tests reveal significant performance disparities for leading models on non-standard dialects. Second, we use this system as a data augmentation technique to improve the dialect robustness of existing systems. Finally, we partner with native speakers of Chicano and Indian English to release new gold-standard variants of the popular CoQA task. To execute the transformation code, run model checkpoints, and download both synthetic and gold-standard dialectal benchmark datasets, see http://value-nlp.org.

📄 PDF Abstract BibTeX arXiv:2212.08011

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationMachine TranslationQuestion AnsweringSemantic ParsingTranslation

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

2026-04-07 · Jason Lucas, Matt Murtagh, Ali Al-Lawati, Uchendu Uchendu 외 arxiv

Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), leaving their robustness to dialectal variation unexplored. We present DIA…

DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models

2025-10-31 · Malik H. Altakrori, Nizar Habash, Abed Alhakim Freihat, Younes Samih 외 arxiv

We present DialectalArabicMMLU, a new benchmark for evaluating the performance of large language models (LLMs) across Arabic dialects. While recently developed Arabic and multilingual benchmarks have advanced LLM evaluat…

DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation

2026-07-08 · Jordan Painter, Dipankar Srirag, Adarsh Kappiyath, Diptesh Kanojia 외 arxiv

Large language models increasingly \emph{understand} dialectal English, yet still \emph{produce} only standard, US-leaning English, leaving dialectal generation, the harder half of the problem, largely unaddressed. We in…

Continual Pretraining

VALUE: Understanding Dialect Disparity in NLU

2022-04-06 · ACL 2022 5 · Caleb Ziems, Jiaao Chen, Camille Harris, Jessica Anderson 외

English Natural Language Understanding (NLU) systems have achieved great performances and even outperformed humans on benchmarks like GLUE and SuperGLUE. However, these benchmarks contain only textbook Standard American …

Linguistic AcceptabilityNatural Language Understanding

Which English Do LLMs Prefer? Triangulating Structural Bias Towards American English in Foundation Models

2026-04-05 · Mir Tafseer Nayeem, Davood Rafiei arxiv

Large language models (LLMs) are increasingly deployed in high-stakes domains, yet they expose only limited language settings, most notably "English (US)," despite the global diversity and colonial history of English. Th…