paper-with-me

홈 › Papers

Probing Cultural Signals in Large Language Models through Author Profiling

2026-03-17 · Valentin Lafargue, Ariel Guerra-Adames, Emmanuelle Claeys, Elouan Vuichard, Jean-Michel Loubes arxiv

Large language models (LLMs) are increasingly deployed in applications with societal impact, raising concerns about the cultural biases they encode. We probe these representations by evaluating whether LLMs can perform author profiling from song lyrics in a zero-shot setting, inferring singers' gender and ethnicity without task-specific fine-tuning. Across several open-source models evaluated on more than 10,000 lyrics, we find that LLMs achieve non-trivial profiling performance but demonstrate systematic cultural alignment: most models default toward North American ethnicity, while DeepSeek-1.5B aligns more strongly with Asian ethnicity. This finding emerges from both the models' prediction distributions and an analysis of their generated rationales. To quantify these disparities, we introduce two fairness metrics, Modality Accuracy Divergence (MAD) and Recall Divergence (RD), and show that Ministral-8B displays the strongest ethnicity bias among the evaluated models, whereas Gemma-12B shows the most balanced behavior. Our code is available on GitHub and results on HuggingFace.

📄 PDF Abstract BibTeX arXiv:2603.16749

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs

2025-11-15 · Truong Vo, Sanmi Koyejo arxiv

Large language models (LLMs) are increasingly deployed in culturally diverse environments, yet existing evaluations of cultural competence remain limited. Existing methods focus on de-contextualized correctness or forced…

Does Mapo Tofu Contain Coffee? Probing LLMs for Food-related Cultural Knowledge

2024-04-10 · Li Zhou, Taelin Karidi, Wanlong Liu, Nicolas Garneau 외

Recent studies have highlighted the presence of cultural biases in Large Language Models (LLMs), yet often lack a robust methodology to dissect these phenomena comprehensively. Our work aims to bridge this gap by delving…

Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation

2025-11-21 · Chuancheng Shi, Shangze Li, Shiming Guo, Simiao Xie 외 arxiv

Multilingual text-to-image (T2I) models have advanced rapidly in terms of visual realism and semantic alignment, and are now widely utilized. Yet outputs vary across cultural contexts: because language carries cultural c…

Text-to-Image Generation

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

2026-06-09 · Trung Duc Anh Dang, Tung Kieu, Sarah Masud arxiv

Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cultural alignment typically rely on direct prompting with survey-style…

Disentangling Language and Culture for Evaluating Multilingual Large Language Models

2025-05-30 · Jiahao Ying, Wei Tang, Yiran Zhao, Yixin Cao 외

This paper introduces a Dual Evaluation Framework to comprehensively assess the multilingual capabilities of LLMs. By decomposing the evaluation along the dimensions of linguistic medium and cultural context, this framew…