paper-with-me

Papers

D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation

2024-04-16 · Aida Mostafazadeh Davani, Mark Díaz, Dylan Baker, Vinodkumar Prabhakaran

While human annotations play a crucial role in language technologies, annotator subjectivity has long been overlooked in data collection. Recent studies that have critically examined this issue are often situated in the Western context, and solely document differences across age, gender, or racial groups. As a result, NLP research on subjectivity have overlooked the fact that individuals within demographic groups may hold diverse values, which can influence their perceptions beyond their group norms. To effectively incorporate these considerations into NLP pipelines, we need datasets with extensive parallel annotations from various social and cultural groups. In this paper we introduce the \dataset dataset: a large-scale cross-cultural dataset of parallel annotations for offensive language in over 4.5K sentences annotated by a pool of over 4k annotators, balanced across gender and age, from across 21 countries, representing eight geo-cultural regions. The dataset contains annotators' moral values captured along six moral foundations: care, equality, proportionality, authority, loyalty, and purity. Our analyses reveal substantial regional variations in annotators' perceptions that are shaped by individual moral values, offering crucial insights for building pluralistic, culturally sensitive NLP models.

📄 PDF Abstract BibTeX arXiv:2404.10857

Code (0)

등록된 구현이 없습니다.

Tasks

4k

Similar Papers 제목 키워드 기반

Isolating Culture Neurons in Multilingual Large Language Models

2025-08-04 · Danial Namazifard, Lukas Galke Poech arxiv

Language and culture are deeply intertwined, yet it has been unclear how and where multilingual large language models encode culture. Here, we build on an established methodology for identifying language-specific neurons…

Who Laughs with Whom? Disentangling Influential Factors in Humor Preferences across User Clusters and LLMs

2026-01-06 · Soichiro Murakami, Hidetaka Kamigaito, Hiroya Takamura, Manabu Okumura arxiv

Humor preferences vary widely across individuals and cultures, complicating the evaluation of humor using large language models (LLMs). In this study, we model heterogeneity in humor preferences in Oogiri, a Japanese cre…

DEBATE: A Dataset for Disentangling Textual Ambiguity in Mandarin Through Speech

2025-06-09 · Haotian Guo, Jing Han, Yongfeng Tu, Shihao Gao 외

Despite extensive research on textual and visual disambiguation, disambiguation through speech (DTS) remains underexplored. This is largely due to the lack of high-quality datasets that pair spoken sentences with richly …

Probing Pre-Trained Language Models for Cross-Cultural Differences in Values

2022-03-25 · Arnav Arora, Lucie-Aimée Kaffee, Isabelle Augenstein

Language embeds information about social, cultural, and political values people hold. Prior work has explored social and potentially harmful biases encoded in Pre-Trained Language models (PTLMs). However, there has been …

CultureScore: Evaluating Cultural Faithfulness in Video Generation Models

2026-06-05 · Anku Rani, Wei Dai, Shravan Nayak, Pattie Maes 외 arxiv

As video generation models like Veo 3.1 and LTX-2 advance, their ability to accurately represent diverse global cultures remains a critical yet understudied frontier. Current metrics, such as VideoScore, only measure vis…

Video Generation