paper-with-me

Papers

Unsupervised Discovery of Implicit Gender Bias

2020-04-17 · EMNLP 2020 11 · Anjalie Field, Yulia Tsvetkov

Despite their prevalence in society, social biases are difficult to identify, primarily because human judgements in this domain can be unreliable. We take an unsupervised approach to identifying gender bias against women at a comment level and present a model that can surface text likely to contain bias. Our main challenge is forcing the model to focus on signs of implicit bias, rather than other artifacts in the data. Thus, our methodology involves reducing the influence of confounds through propensity matching and adversarial learning. Our analysis shows how biased comments directed towards female politicians contain mixed criticisms, while comments directed towards other female public figures focus on appearance and sexualization. Ultimately, our work offers a way to capture subtle biases in various domains without relying on subjective human judgements.

📄 PDF Abstract BibTeX arXiv:2004.08361

Code (1)

anjalief/unsupervised_gender_bias 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Semi-Supervised Topic Modeling for Gender Bias Discovery in English and Swedish

2020-12-01 · GeBNLP (COLING) 2020 12 · Hannah Devinney, Jenny Björklund, Henrik Björklund

Gender bias has been identified in many models for Natural Language Processing, stemming from implicit biases in the text corpora used to train the models. Such corpora are too large to closely analyze for biased or ster…

ArticlesTopic Models

Image Representations Learned With Unsupervised Pre-Training Contain Human-like Biases

2020-10-28 · Ryan Steed, Aylin Caliskan

Recent advances in machine learning leverage massive datasets of unlabeled images from the web to learn general-purpose image representations for tasks from image classification to face recognition. But do unsupervised c…

BIG-bench Machine LearningFace Recognitionimage-classificationImage Classification+1

Probing Explicit and Implicit Gender Bias through LLM Conditional Text Generation

2023-11-01 · Xiangjue Dong, Yibo Wang, Philip S. Yu, James Caverlee

Large Language Models (LLMs) can generate biased and toxic responses. Yet most prior work on LLM gender bias evaluation requires predefined gender-related phrases or gender stereotypes, which are challenging to be compre…

Conditional Text GenerationFairnessText Generation

Transcending the "Male Code": Implicit Masculine Biases in NLP Contexts

2023-04-22 · Katie Seaborn, Shruti Chandra, Thibault Fabre

Critical scholarship has elevated the problem of gender bias in data sets used to train virtual assistants (VAs). Most work has focused on explicit biases in language, especially against women, girls, femme-identifying p…

Word Embeddings

Implicit Bias in LLMs for Transgender Populations

2026-02-02 · Micaela Hirsch, Marina Elichiry, Blas Radi, Tamara Quiroga 외 arxiv

Large language models (LLMs) have been shown to exhibit biases against LGBTQ+ populations. While safety training may lessen explicit expressions of bias, previous work has shown that implicit stereotype-driven associatio…