paper-with-me

홈 › Papers

Defining Cultural Capabilities for AI Evaluation: A Taxonomy Grounded in Intercultural Communication Theory

2026-05-15 · Isar Nejadgholi, Masoud Kianpour, Krishnapriya Vishnubhotla, Maryam Molamohamadi arxiv

Tremendous efforts have been put into evaluating the inclusivity and effectiveness of AI systems across cultures. However, the cultural capabilities considered in much of the literature remain vaguely defined, are referred to using interchangeable terminology, and are typically limited to recalling accurate information about various demographics, regions, and nationalities. To address this construct ambiguity, we draw from Intercultural Communication scholarship and propose a three-level taxonomy of AI-relevant cultural capabilities: Cultural Awareness answers "Does the model know?", Cultural Sensitivity answers "How does it frame its knowledge?", and Cultural Competence answers "Can it adapt as the interaction evolves?". Beyond conceptual clarification, we position this taxonomy as a practical tool for improving the validity and interpretability of AI evaluation in real-world, multicultural settings. Without such construct clarity, evaluation results risk overstating model capabilities and may lead to inappropriate deployment decisions in culturally sensitive contexts.

📄 PDF Abstract BibTeX arXiv:2605.15990

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Words to Worlds: Benchmarking Cross-Cultural Cultural Understanding in Machine Translation

2026-03-18 · Bangju Han, Yingqi Wang, Huang Qing, Tiyuan Li 외 arxiv

Culture-expressions, such as idioms, slang, and culture-specific items (CSIs), are pervasive in natural language and encode meanings that go beyond literal linguistic form. Accurately translating such expressions remains…

Machine Translation

Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach

2025-08-06 · Kezia Oketch, John P. Lalor, Ahmed Abbasi arxiv

We introduce the first taxonomy-guided evaluation of Swahili NLP, addressing gaps in sociolinguistic diversity. Drawing on health-related psychometric tasks, we collect a dataset of 2,170 free-text responses from Kenyan …

Beyond English and Evasion: A Human-Annotated Multi-Domain Benchmark for High-Stakes LLM Safety Evaluation in Chinese

2026-05-28 · Wajdi Zaghouani, Kholoud K. Aldous, Yicheng Gao arxiv

When Large Language Models (LLMs) are deployed in Chinese-language settings, a troubling pattern emerges: safety systems that work well in English break down. These systems struggle to cross linguistic and cultural bound…

AlignCultura: Towards Culturally Aligned Large Language Models?

2026-04-21 · Gautam Siddharth Kashyap, Mark Dras, Usman Naseem arxiv

Cultural alignment in Large Language Models (LLMs) is essential for producing contextually aware, respectful, and trustworthy outputs. Without it, models risk generating stereotyped, insensitive, or misleading responses …

Response Generation

Identification of Stone Deterioration Patterns with Large Multimodal Models

2024-06-05 · Daniele Corradetti, Jose Delgado Rodrigues

The conservation of stone-based cultural heritage sites is a critical concern for preserving cultural and historical landmarks. With the advent of Large Multimodal Models, as GPT-4omni (OpenAI), Claude 3 Opus (Anthropic)…

Image Classification