paper-with-me

Papers

De-biasing "bias" measurement

2022-05-11 · Kristian Lum, Yunfeng Zhang, Amanda Bower

When a model's performance differs across socially or culturally relevant groups--like race, gender, or the intersections of many such groups--it is often called "biased." While much of the work in algorithmic fairness over the last several years has focused on developing various definitions of model fairness (the absence of group-wise model performance disparities) and eliminating such "bias," much less work has gone into rigorously measuring it. In practice, it important to have high quality, human digestible measures of model performance disparities and associated uncertainty quantification about them that can serve as inputs into multi-faceted decision-making processes. In this paper, we show both mathematically and through simulation that many of the metrics used to measure group-wise model performance disparities are themselves statistically biased estimators of the underlying quantities they purport to represent. We argue that this can cause misleading conclusions about the relative group-wise model performance disparities along different dimensions, especially in cases where some sensitive variables consist of categories with few members. We propose the "double-corrected" variance estimator, which provides unbiased estimates and uncertainty quantification of the variance of model performance across groups. It is conceptually simple and easily implementable without statistical software package or numerical optimization. We demonstrate the utility of this approach through simulation and show on a real dataset that while statistically biased estimators of group-wise model performance disparities indicate statistically significant differences, when accounting for statistical bias in the estimator, the estimated between-group disparities are no longer statistically significant.

📄 PDF Abstract BibTeX arXiv:2205.05770

Code (1)

twitter-research/double-corrected-variance-estimator 공식 구현

Tasks

Decision MakingFairnessUncertainty Quantification

Similar Papers 제목 키워드 기반

Information-Theoretic Bias Reduction via Causal View of Spurious Correlation

2022-01-10 · Seonguk Seo, Joon-Young Lee, Bohyung Han

We propose an information-theoretic bias measurement technique through a causal interpretation of spurious correlation, which is effective to identify the feature-level algorithmic bias by taking advantage of conditional…

Face RecognitionFairness

RedditBias: A Real-World Resource for Bias Evaluation and Debiasing of Conversational Language Models

2021-06-07 · ACL 2021 5 · Soumya Barikeri, Anne Lauscher, Ivan Vulić, Goran Glavaš

Text representation models are prone to exhibit a range of societal biases, reflecting the non-controlled and biased nature of the underlying pretraining data, which consequently leads to severe ethical issues and even b…

Conversational Response GenerationResponse Generation

A Prompt Array Keeps the Bias Away: Debiasing Vision-Language Models with Adversarial Learning

2022-03-22 · Hugo Berg, Siobhan Mackenzie Hall, Yash Bhalgat, Wonsuk Yang 외

Vision-language models can encode societal biases and stereotypes, but there are challenges to measuring and mitigating these multimodal harms due to lacking measurement robustness and feature degradation. To address the…

Optimal biasing and physical limits of DVS event noise

2023-04-08 · Rui Graca, Brian Mcreynolds, Tobi Delbruck

Under dim lighting conditions, the output of Dynamic Vision Sensor (DVS) event cameras is strongly affected by noise. Photon and electron shot-noise cause a high rate of non-informative events that reduce Signal to Noise…

Evaluating and Mitigating Social Bias for Large Language Models in Open-ended Settings

2024-12-09 · Zhao Liu, Tian Xie, Xueru Zhang

Current social bias benchmarks for Large Language Models (LLMs) primarily rely on pre-defined question formats like multiple-choice, limiting their ability to reflect the complexity and open-ended nature of real-world in…

Multiple-choice