paper-with-me

Papers

QueerBench: Quantifying Discrimination in Language Models Toward Queer Identities

2024-06-18 · Mae Sosto, Alberto Barrón-Cedeño

With the increasing role of Natural Language Processing (NLP) in various applications, challenges concerning bias and stereotype perpetuation are accentuated, which often leads to hate speech and harm. Despite existing studies on sexism and misogyny, issues like homophobia and transphobia remain underexplored and often adopt binary perspectives, putting the safety of LGBTQIA+ individuals at high risk in online spaces. In this paper, we assess the potential harm caused by sentence completions generated by English large language models (LLMs) concerning LGBTQIA+ individuals. This is achieved using QueerBench, our new assessment framework, which employs a template-based approach and a Masked Language Modeling (MLM) task. The analysis indicates that large language models tend to exhibit discriminatory behaviour more frequently towards individuals within the LGBTQIA+ community, reaching a difference gap of 7.2% in the QueerBench score of harmfulness.

📄 PDF Abstract BibTeX arXiv:2406.12399

Code (1)

maesosto/queerbench 공식 구현

Tasks

Language ModelingLanguage ModellingMasked Language ModelingSentence

Similar Papers 제목 키워드 기반

PRIDE -- Parameter-Efficient Reduction of Identity Discrimination for Equality in LLMs

2025-07-18 · Maluna Menke, Thilo Hagendorff arxiv

Large Language Models (LLMs) frequently reproduce the gender- and sexual-identity prejudices embedded in their training corpora, leading to outputs that marginalize LGBTQIA+ users. Hence, reducing such biases is of great…

parameter-efficient fine-tuning

Un-Straightening Generative AI: How Queer Artists Surface and Challenge the Normativity of Generative AI Models

2025-03-12 · Jordan Taylor, Joel Mire, Franchesca Spektor, Alicia DeVrio 외

Queer people are often discussed as targets of bias, harm, or discrimination in research on generative AI. However, the specific ways that queer people engage with generative AI, and thus possible uses that support queer…

Queer People are People First: Deconstructing Sexual Identity Stereotypes in Large Language Models

2023-06-30 · Harnoor Dhingra, Preetiha Jayashanker, Sayali Moghe, Emma Strubell

Large Language Models (LLMs) are trained primarily on minimally processed web text, which exhibits the same wide range of social biases held by the humans who created that content. Consequently, text generated by LLMs ca…

Sentence

SLAyiNG: A Diverse and Community-validated Dataset of Queer Slang

2025-09-22 · Leonor Veloso, Lea Hirlimann, Philipp Wicke, Hinrich Schütze arxiv

Queer vernacular is rarely studied in NLP, despite advancements in resources and evaluation for other sociolects and informal language. Because of this, NLP systems often process queer language incorrectly, e.g., they mi…

Mitigating Bias in Queer Representation within Large Language Models: A Collaborative Agent Approach

2024-11-12 · Tianyi Huang, Arya Somasundaram

Large Language Models (LLMs) often perpetuate biases in pronoun usage, leading to misrepresentation or exclusion of queer individuals. This paper addresses the specific problem of biased pronoun usage in LLM outputs, par…

Bias DetectionFairnessMulti-agent Integration