paper-with-me

홈 › Papers

Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings

2025-04-04 · Kaustubh Shivshankar Shejole, Pushpak Bhattacharyya

Stereotypes are known to be highly pernicious, making their detection critically important. However, current research predominantly focuses on detecting and evaluating stereotypical biases in LLMs, leaving the study of stereotypes in its early stages. Many studies have failed to clearly distinguish between stereotypes and stereotypical biases, which has significantly slowed progress in advancing research in this area. Stereotype and anti-stereotype detection is a problem that requires knowledge of society; hence, it is one of the most difficult areas in Responsible AI. This work investigates this task, where we propose a four-tuple definition and provide precise terminology distinguishing stereotype, anti-stereotype, stereotypical bias, and bias, offering valuable insights into their various aspects. In this paper, we propose StereoDetect, a high-quality benchmarking dataset curated for this task by optimally utilizing current datasets such as StereoSet and WinoQueer, involving a manual verification process and the transfer of semantic information. We demonstrate that language models for reasoning with fewer than 10B parameters often get confused when detecting anti-stereotypes. We also demonstrate the critical importance of well-curated datasets by comparing our model with other current models for stereotype detection. The dataset and code is available at https://github.com/KaustubhShejole/StereoDetect.

📄 PDF Abstract BibTeX arXiv:2504.03352

Code (1)

kaustubhshejole/stereodetect 공식 구현

Tasks

Benchmarking

Similar Papers 제목 키워드 기반

Quantifying Stereotypes in Language

2024-01-28 · Yang Liu

A stereotype is a generalized perception of a specific group of humans. It is often potentially encoded in human language, which is more common in texts on social issues. Previous works simply define a sentence as stereo…

Sentence

Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model

2021-06-04 · ACL 2021 5 · Kathleen C. Fraser, Isar Nejadgholi, Svetlana Kiritchenko

Stereotypical language expresses widely-held beliefs about different social categories. Many stereotypes are overtly negative, while others may appear positive on the surface, but still lead to negative consequences. In …

Extracting Age-Related Stereotypes from Social Media Texts

2022-06-01 · LREC 2022 6 · Kathleen C. Fraser, Svetlana Kiritchenko, Isar Nejadgholi

Age-related stereotypes are pervasive in our society, and yet have been under-studied in the NLP community. Here, we present a method for extracting age-related stereotypes from Twitter data, generating a corpus of 300,0…

Language Agents for Detecting Implicit Stereotypes in Text-to-image Models at Scale

2023-10-18 · Qichao Wang, Tian Bian, Yian Yin, Tingyang Xu 외

The recent surge in the research of diffusion models has accelerated the adoption of text-to-image models in various Artificial Intelligence Generated Content (AIGC) commercial products. While these exceptional AIGC prod…

Hate Speech Classifiers Learn Human-Like Social Stereotypes

2021-10-28 · Aida Mostafazadeh Davani, Mohammad Atari, Brendan Kennedy, Morteza Dehghani

Social stereotypes negatively impact individuals' judgements about different groups and may have a critical role in how people understand language directed toward minority social groups. Here, we assess the role of socia…

Fairness