paper-with-me

홈 › Papers

DIWALI: Diversity and Inclusivity aWare cuLture specific Items for India: Dataset and Assessment of LLMs for Cultural Text Adaptation in Indian Context

2025-09-22 · Pramit Sahoo, Maharaj Brahma, Maunendra Sankar Desarkar arxiv

Large language models (LLMs) are widely used in various tasks and applications. However, despite their wide capabilities, they are shown to lack cultural alignment \citep{ryan-etal-2024-unintended, alkhamissi-etal-2024-investigating} and produce biased generations \cite{naous-etal-2024-beer} due to a lack of cultural knowledge and competence. Evaluation of LLMs for cultural awareness and alignment is particularly challenging due to the lack of proper evaluation metrics and unavailability of culturally grounded datasets representing the vast complexity of cultures at the regional and sub-regional levels. Existing datasets for culture specific items (CSIs) focus primarily on concepts at the regional level and may contain false positives. To address this issue, we introduce a novel CSI dataset for Indian culture, belonging to 17 cultural facets. The dataset comprises ~8k cultural concepts from 36 sub-regions. To measure the cultural competence of LLMs on a cultural text adaptation task, we evaluate the adaptations using the CSIs created, LLM as Judge, and human evaluations from diverse socio-demographic region. Furthermore, we perform quantitative analysis demonstrating selective sub-regional coverage and surface-level adaptations across all considered LLMs. Our dataset is available here: https://huggingface.co/datasets/nlip/DIWALI, project webpage https://nlip-lab.github.io/nlip/publications/diwali/, and our codebase with model outputs can be found here: https://github.com/pramitsahoo/culture-evaluation

📄 PDF Abstract BibTeX arXiv:2509.17399

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis

2025-05-26 · Ruixiang Feng, Shen Gao, Xiuying Chen, Lisi Chen 외

Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks, yet they often exhibit a specific cultural biases, neglecting the values and linguistic diversity of low-resource regions. This…

DiversityOpen-Ended Question AnsweringQuestion Answering

From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models

2024-06-28 · Mehar Bhatia, Sahithya Ravi, Aditya Chinchure, EunJeong Hwang 외

Despite recent advancements in vision-language models, their performance remains suboptimal on images from non-western cultures due to underrepresentation in training datasets. Various benchmarks have been proposed to te…

DiversityRetrievalVisual Grounding

Artificial Intelligence for Inclusive Engineering Education: Advancing Equality, Diversity, and Ethical Leadership

2026-01-24 · Mona G. Ibrahim, Riham Hilal arxiv

AI technology development has transformed the field of engineering education with its adaptivity-driven, data-based, and ethical-led learning platforms that promote equity, diversity, and inclusivity. But with so much pr…

Social Bias in Multilingual Language Models: A Survey

2025-08-27 · Lance Calvin Lim Gamboa, Yue Feng, Mark Lee arxiv

Pretrained multilingual models exhibit the same social bias as models processing English texts. This systematic review analyzes emerging research that extends bias evaluation and mitigation approaches into multilingual a…

Isolating Culture Neurons in Multilingual Large Language Models

2025-08-04 · Danial Namazifard, Lukas Galke Poech arxiv

Language and culture are deeply intertwined, yet it has been unclear how and where multilingual large language models encode culture. Here, we build on an established methodology for identifying language-specific neurons…