paper-with-me

홈 › Papers

CultureScope: A Dimensional Lens for Probing Cultural Understanding in LLMs

2025-09-19 · Jinghao Zhang, Sihang Jiang, Shiwei Guo, Shisong Chen, Yanghua Xiao, Hongwei Feng, Jiaqing Liang, Minggui HE, Shimin Tao, Hongxia Ma arxiv

As large language models (LLMs) are increasingly deployed in diverse cultural environments, evaluating their cultural understanding capability has become essential for ensuring trustworthy and culturally aligned applications. However, most existing benchmarks lack comprehensiveness and are challenging to scale and adapt across different cultural contexts, because their frameworks often lack guidance from well-established cultural theories and tend to rely on expert-driven manual annotations. To address these issues, we propose CultureScope, the most comprehensive evaluation framework to date for assessing cultural understanding in LLMs. Inspired by the cultural iceberg theory, we design a novel dimensional schema for cultural knowledge classification, comprising 3 layers and 140 dimensions, which guides the automated construction of culture-specific knowledge bases and corresponding evaluation datasets for any given languages and cultures. Experimental results demonstrate that our method can effectively evaluate cultural understanding. They also reveal that existing large language models lack comprehensive cultural competence, and merely incorporating multilingual data does not necessarily enhance cultural understanding. All code and data files are available at https://github.com/HoganZinger/Culture

📄 PDF Abstract BibTeX arXiv:2509.16188

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CULEMO: Cultural Lenses on Emotion -- Benchmarking LLMs for Cross-Cultural Emotion Understanding

2025-03-12 · Tadesse Destaw Belay, Ahmed Haj Ahmed, Alvin Grissom II, Iqra Ameer 외

NLP research has increasingly focused on subjective tasks such as emotion analysis. However, existing emotion benchmarks suffer from two major shortcomings: (1) they largely rely on keyword-based emotion recognition, ove…

BenchmarkingEmotion RecognitionSentiment Analysis

Does Mapo Tofu Contain Coffee? Probing LLMs for Food-related Cultural Knowledge

2024-04-10 · Li Zhou, Taelin Karidi, Wanlong Liu, Nicolas Garneau 외

Recent studies have highlighted the presence of cultural biases in Large Language Models (LLMs), yet often lack a robust methodology to dissect these phenomena comprehensively. Our work aims to bridge this gap by delving…

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

2024-12-17 · Yangfan Ye, Xiaocheng Feng, Xiachong Feng, Libo Qin 외

Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their English-centric pre-training data. In this paper, we introduce and investi…

CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs

2025-11-15 · Truong Vo, Sanmi Koyejo arxiv

Large language models (LLMs) are increasingly deployed in culturally diverse environments, yet existing evaluations of cultural competence remain limited. Existing methods focus on de-contextualized correctness or forced…

Probing Cultural Awareness in LLMs: A Case Study of Cross-Culture Aesthetic Stylistics

2026-05-26 · Jiashuo Wang, Fenggang Yu, Jian Wang, Chak Tou Leong 외 arxiv

Large Language Models (LLMs) are increasingly deployed in diverse cultural contexts, yet their ability to master aesthetic stylistics, i.e., the strategic use of language to evoke cultural resonance, remains underexplore…