paper-with-me

홈 › Papers

Representational Harms in LLM-Generated Narratives Against Global Majority Nationalities

2026-04-24 · Ilana Nguyen, Harini Suresh, Thema Monroe-White, Evan Shieh arxiv

Large language models (LLMs) are increasingly used for text generation tasks from everyday use to high-stakes enterprise and government applications, including simulated interviews with asylum seekers. While many works highlight the new potential applications of LLMs, there are risks of LLMs encoding and perpetuating harmful biases about non-dominant communities across the globe. To better evaluate and mitigate such harms, more research examining how LLMs portray diverse individuals is needed. In this work, we study how national origin identities are portrayed by widely-adopted LLMs in response to open-ended narrative generation prompts. Our findings demonstrate the presence of persistent representational harms by national origin, including harmful stereotypes, erasure, and one-dimensional portrayals of Global Majority identities. Minoritized national identities are simultaneously underrepresented in power-neutral stories and overrepresented in subordinated character portrayals, which are over fifty times more likely to appear than dominant portrayals. The degree of harm is amplified when US nationality cues (e.g., ``American'') are present in input prompts. Notably, we find that the harms we identify cannot be explained away via sycophancy, as US-centric biases persist even when replacing US nationality cues with non-US national identities in the prompts. Based on our findings, we call for further exploration of cultural harms in LLMs through methodologies that center Global Majority perspectives and challenge the uncritical adoption of US-based LLMs for the classification, surveillance, and misrepresentation of the majority of our planet.

📄 PDF Abstract BibTeX arXiv:2604.22749

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Words of Wisdom: Representational Harms in Learning From AI Communication

2021-11-16 · Amanda Buddemeyer, Erin Walker, Malihe Alikhani

Many educational technologies use artificial intelligence (AI) that presents generated or produced language to the learner. We contend that all language, including all AI communication, encodes information about the iden…

DiversityQuestion GenerationQuestion-Generation

Beyond Behaviorist Representational Harms: A Plan for Measurement and Mitigation

2024-01-25 · Jennifer Chien, David Danks

Algorithmic harms are commonly categorized as either allocative or representational. This study specifically addresses the latter, focusing on an examination of current definitions of representational harms to discern wh…

Fairness

Taxonomizing Representational Harms using Speech Act Theory

2025-04-01 · Emily Corvi, Hannah Washington, Stefanie Reed, Chad Atalla 외

Representational harms are widely recognized among fairness-related harms caused by generative language systems. However, their definitions are commonly under-specified. We make a theoretical contribution to the specific…

Fairnessvalid

The Psychosocial Impacts of Generative AI Harms

2024-05-02 · Faye-Marie Vassel, Evan Shieh, Cassidy R. Sugimoto, Thema Monroe-White

The rapid emergence of generative Language Models (LMs) has led to growing concern about the impacts that their unexamined adoption may have on the social well-being of diverse user groups. Meanwhile, LMs are increasingl…

An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models

2023-01-22 · Saghar Hosseini, Hamid Palangi, Ahmed Hassan Awadallah

Large-scale Pre-Trained Language Models (PTLMs) capture knowledge from massive human-written data which contains latent societal biases and toxic contents. In this paper, we leverage the primary task of PTLMs, i.e., lang…

Language ModelingLanguage Modelling