paper-with-me

Papers

Generative Language Models Exhibit Social Identity Biases

2023-10-24 · Tiancheng Hu, Yara Kyrychenko, Steve Rathje, Nigel Collier, Sander van der Linden, Jon Roozenbeek

The surge in popularity of large language models has given rise to concerns about biases that these models could learn from humans. We investigate whether ingroup solidarity and outgroup hostility, fundamental social identity biases known from social psychology, are present in 56 large language models. We find that almost all foundational language models and some instruction fine-tuned models exhibit clear ingroup-positive and outgroup-negative associations when prompted to complete sentences (e.g., "We are..."). Our findings suggest that modern language models exhibit fundamental social identity biases to a similar degree as humans, both in the lab and in real-world conversations with LLMs, and that curating training data and instruction fine-tuning can mitigate such biases. Our results have practical implications for creating less biased large-language models and further underscore the need for more research into user interactions with LLMs to prevent potential bias reinforcement in humans.

📄 PDF Abstract BibTeX arXiv:2310.15819

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating Biased Attitude Associations of Language Models in an Intersectional Context

2023-07-07 · Shiva Omrani Sabbaghi, Robert Wolfe, Aylin Caliskan

Language models are trained on large-scale corpora that embed implicit biases documented in psychology. Valence associations (pleasantness/unpleasantness) of social groups determine the biased attitudes towards groups an…

SentenceWord Embeddings

Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups

2025-10-08 · Geng Liu, Feng Li, Junjie Mu, Mengxiao Zhu 외 arxiv

Large language models (LLMs) are increasingly deployed in user-facing applications, raising concerns that they may reflect and amplify social biases. We investigate social identity biases in Chinese LLMs using Mandarin-s…

Towards Socially Responsible AI: Cognitive Bias-Aware Multi-Objective Learning

2020-05-14 · Procheta Sen, Debasis Ganguly

Human society had a long history of suffering from cognitive biases leading to social prejudices and mass injustice. The prevalent existence of cognitive biases in large volumes of historical data can pose a threat of be…

Abusive Language

Laissez-Faire Harms: Algorithmic Biases in Generative Language Models

2024-04-11 · Evan Shieh, Faye-Marie Vassel, Cassidy Sugimoto, Thema Monroe-White

The rapid deployment of generative language models (LMs) has raised concerns about social biases affecting the well-being of diverse consumers. The extant literature on generative LMs has primarily examined bias via expl…

Analyzing Social Biases in Japanese Large Language Models

2024-06-04 · Hitomi Yanaka, Namgi Han, Ryoma Kumon, Jie Lu 외

With the development of Large Language Models (LLMs), social biases in the LLMs have become a crucial issue. While various benchmarks for social biases have been provided across languages, the extent to which Japanese LL…

Question Answering