paper-with-me

홈 › Papers

LLMs Infer Cultural Context but Fail to Apply It When Responding

2026-06-16 · Yisong Miao, Jian Zhu, Vered Shwartz arxiv

Recent work has shown that LLMs overrepresent dominant cultures, particularly Western ones, while marginalizing others. We investigate whether this affects models' ability to generate culturally adapted responses by evaluating their use of local measurement units based on the user's perceived cultural background. We introduce Cultural and Pragmatic Response Inference (CAPRI), a dataset of conversations with varying levels of cultural cues. Experiments with state-of-the-art LLMs show that models can infer cultural background and recall relevant conventions, but often fail to utilize the information to adapt their answers to the relevant cultural conventions, unless explicitly prompted to perform the tasks sequentially. We further evaluate adaptation to the interpretation of time and quantity expressions, two subjective language grounding dimensions that are affected by culture. We find that models increasingly adapt their answers as cultural cues accumulate, but their priors are not culture-neutral, sometimes aligning with the model's country of origin. Overall, CAPRI provides a resource for future research aimed at narrowing the gap between cultural knowledge and culturally adaptive language generation.

📄 PDF Abstract BibTeX arXiv:2606.17688

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

BanglaSocialBench: A Benchmark for Evaluating Sociopragmatic and Cultural Alignment of LLMs in Bangladeshi Social Interaction

2026-03-16 · Tanvir Ahmed Sijan, S. M Golam Rifat, Pankaj Chowdhury Partha, Md. Tanjeed Islam 외 arxiv

Large Language Models have demonstrated strong multilingual fluency, yet fluency alone does not guarantee socially appropriate language use. In high-context languages, communicative competence requires sensitivity to soc…

Cultural Bias in Large Language Models: Evaluating AI Agents through Moral Questionnaires

2025-07-14 · Simon Münker arxiv

Are AI systems truly representing human values, or merely averaging across them? Our study suggests a concerning reality: Large Language Models (LLMs) fail to represent diverse cultural moral frameworks despite their lin…

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

2026-06-09 · Trung Duc Anh Dang, Tung Kieu, Sarah Masud arxiv

Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cultural alignment typically rely on direct prompting with survey-style…

Can LLMs Grasp Implicit Cultural Values? Benchmarking LLMs' Metacognitive Cultural Intelligence with CQ-Bench

2025-04-01 · Ziyi Liu, Priyanka Dey, Zhenyu Zhao, Jen-tse Huang 외

Cultural Intelligence (CQ) refers to the ability to understand unfamiliar cultural contexts-a crucial skill for large language models (LLMs) to effectively engage with globally diverse users. While existing research ofte…

Benchmarking

When English Rewrites Local Knowledge: Global Narrative Dominance in Large Language Models

2026-05-28 · Md Arid Hasan, Ruwad Naswan, Farhan Samir, Sharifa Sultana 외 arxiv

Large language models (LLMs) are widely used as cross-lingual knowledge interfaces. However, culturally grounded questions often reflect globally dominant narratives rather than local contexts. We study this failure mode…