paper-with-me

Papers

Measuring Machine Learning Harms from Stereotypes Requires Understanding Who Is Harmed by Which Errors in What Ways

2024-02-06 · Angelina Wang, Xuechunzi Bai, Solon Barocas, Su Lin Blodgett

As machine learning applications proliferate, we need an understanding of their potential for harm. However, current fairness metrics are rarely grounded in human psychological experiences of harm. Drawing on the social psychology of stereotypes, we use a case study of gender stereotypes in image search to examine how people react to machine learning errors. First, we use survey studies to show that not all machine learning errors reflect stereotypes nor are equally harmful. Then, in experimental studies we randomly expose participants to stereotype-reinforcing, -violating, and -neutral machine learning errors. We find stereotype-reinforcing errors induce more experientially (i.e., subjectively) harmful experiences, while having minimal changes to cognitive beliefs, attitudes, or behaviors. This experiential harm impacts women more than men. However, certain stereotype-violating errors are more experientially harmful for men, potentially due to perceived threats to masculinity. We conclude that harm cannot be the sole guide in fairness mitigation, and propose a nuanced perspective depending on who is experiencing what harm and why.

📄 PDF Abstract BibTeX arXiv:2402.04420

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessImage Retrieval

Similar Papers 제목 키워드 기반

Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

2022-06-23 · NAACL 2022 7 · Yang Trista Cao, Anna Sotnikova, Hal Daumé III, Rachel Rudinger 외

NLP models trained on text have been shown to reproduce human stereotypes, which can magnify harms to marginalized groups when systems are deployed at scale. We adapt the Agency-Belief-Communion (ABC) stereotype model of…

Sensitivity

Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale

2022-11-07 · Federico Bianchi, Pratyusha Kalluri, Esin Durmus, Faisal Ladhak 외

Machine learning models that convert user-written text descriptions into images are now widely available online and used by millions of users to generate millions of images a day. We investigate the potential for these m…

Image GenerationText to Image GenerationText-to-Image Generation

Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language Technologies

2021-08-27 · EMNLP 2021 11 · Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian 외

Gender is widely discussed in the context of language tasks and when examining the stereotypes propagated by language models. However, current discussions primarily treat gender as binary, which can perpetuate harms such…

Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models

2023-09-15 · Khyati Khandelwal, Manuel Tonneau, Andrew M. Bean, Hannah Rose Kirk 외

Large Language Models (LLMs), now used daily by millions, can encode societal biases, exposing their users to representational harms. A large body of scholarship on LLM bias exists but it predominantly adopts a Western-c…

FairnessLanguage ModellingLarge Language Model

A Prompt Array Keeps the Bias Away: Debiasing Vision-Language Models with Adversarial Learning

2022-03-22 · Hugo Berg, Siobhan Mackenzie Hall, Yash Bhalgat, Wonsuk Yang 외

Vision-language models can encode societal biases and stereotypes, but there are challenges to measuring and mitigating these multimodal harms due to lacking measurement robustness and feature degradation. To address the…