paper-with-me

홈 › Papers

Scalable and Culturally Specific Stereotype Dataset Construction via Human-LLM Collaboration

2026-07-08 · Weicheng Ma, John Guerrerio, Soroush Vosoughi arxiv

Research on stereotypes in large language models (LLMs) has largely focused on English-speaking contexts, due to the lack of datasets in other languages and the high cost of manual annotation in underrepresented cultures. To address this gap, we introduce a cost-efficient human-LLM collaborative annotation framework and apply it to construct EspanStereo, a Spanish-language stereotype dataset spanning multiple Spanish-speaking countries across Europe and Latin America. EspanStereo captures both well-documented stereotypes from prior literature and culturally specific biases absent from English-centric resources. Using LLMs to generate candidate stereotypes and in-culture annotators to validate them, we demonstrate the framework's effectiveness in identifying nuanced, region-specific biases. Our evaluation of Spanish-supporting LLMs using EspanStereo reveals significant variation in stereotypical behavior across countries, highlighting the need for more culturally grounded assessments. Beyond Spanish, our framework is adaptable to other languages and regions, offering a scalable path toward multilingual stereotype benchmarks. This work broadens the scope of stereotype analysis in LLMs and lays the groundwork for comprehensive cross-cultural bias evaluation.

📄 PDF Abstract BibTeX arXiv:2607.07895

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AfriStereo: A Culturally Grounded Dataset for Evaluating Stereotypical Bias in Large Language Models

2025-11-27 · Yann Le Beux, Oluchi Audu, Oche D. Ankeli, Dhananjay Balakrishnan 외 arxiv

Existing AI bias evaluation benchmarks largely reflect Western perspectives, leaving African contexts underrepresented and enabling harmful stereotypes in applications across various domains. To address this gap, we intr…

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

2026-03-24 · Sina Bagheri Nezhad arxiv

Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional quality from multimodal and linguistic signals. In multicultural and mu…

SeeGULL Multilingual: a Dataset of Geo-Culturally Situated Stereotypes

2024-03-08 · Mukul Bhutani, Kevin Robinson, Vinodkumar Prabhakaran, Shachi Dave 외

While generative multilingual models are rapidly being deployed, their safety and fairness evaluations are largely limited to resources collected in English. This is especially problematic for evaluations targeting inher…

Fairness

SESGO: Spanish Evaluation of Stereotypical Generative Outputs

2025-09-03 · Melissa Robles, Catalina Bernal, Denniss Raigoso, Mateo Dulce Rubio arxiv

This paper addresses the critical gap in evaluating bias in multilingual Large Language Models (LLMs), with a specific focus on Spanish language within culturally-aware Latin American contexts. Despite widespread global …

Building Socio-culturally Inclusive Stereotype Resources with Community Engagement

2023-07-20 · NeurIPS 2023 11

With rapid development and deployment of generative language models in global settings, there is an urgent need to also scale our measurements of harm, not just in the number and types of harms covered, but also how well…