Fairness in representation: quantifying stereotyping as a representational harm
While harms of allocation have been increasingly studied as part of the subfield of algorithmic fairness, harms of representation have received considerably less attention. In this paper, we formalize two notions of stereotyping and show how they manifest in later allocative harms within the machine learning pipeline. We also propose mitigation strategies and demonstrate their effectiveness on synthetic datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
BIG-bench Machine LearningFairnessSimilar Papers 제목 키워드 기반
Taxonomizing Representational Harms using Speech Act Theory
Representational harms are widely recognized among fairness-related harms caused by generative language systems. However, their definitions are commonly under-specified. We make a theoretical contribution to the specific…
FairnessvalidTowards Understanding and Mitigating Social Biases in Language Models
As machine learning methods are deployed in real-world settings such as healthcare, legal systems, and social science, it is crucial to recognize how they shape social biases and stereotypes in these sensitive decision-m…
Decision MakingFairnessText GenerationFairDistillation: Mitigating Stereotyping in Language Models
Large pre-trained language models are successfully being used in a variety of tasks, across many languages. With this ever-increasing usage, the risk of harmful side effects also rises, for example by reproducing and rei…
Knowledge DistillationFairPrism: Evaluating Fairness-Related Harms in Text Generation
It is critical to measure and mitigate fairness- related harms caused by AI text generation systems, including stereotyping and demeaning harms. To that end, we introduce FairPrism, a dataset of 5,000 examples of AI-gene…
FairnessText GenerationThe Psychosocial Impacts of Generative AI Harms
The rapid emergence of generative Language Models (LMs) has led to growing concern about the impacts that their unexamined adoption may have on the social well-being of diverse user groups. Meanwhile, LMs are increasingl…