Leveraging Labeled and Unlabeled Data for Consistent Fair Binary Classification
We study the problem of fair binary classification using the notion of Equal Opportunity. It requires the true positive rate to distribute equally across the sensitive groups. Within this setting we show that the fair optimal classifier is obtained by recalibrating the Bayes classifier by a group-dependent threshold. We provide a constructive expression for the threshold. This result motivates us to devise a plug-in classification procedure based on both unlabeled and labeled datasets. While the latter is used to learn the output conditional probability, the former is used for calibration. The overall procedure can be computed in polynomial time and it is shown to be statistically consistent both in terms of the classification error and fairness measure. Finally, we present numerical experiments which indicate that our method is often superior or competitive with the state-of-the-art methods on benchmark datasets.
Code (1)
Tasks
Binary ClassificationClassificationFairnessGeneral ClassificationSimilar Papers 제목 키워드 기반
Fairness-aware Model-agnostic Positive and Unlabeled Learning
With the increasing application of machine learning in high-stake decision-making problems, potential algorithmic bias towards people from certain social groups poses negative impacts on individuals and our society at la…
Binary ClassificationDecision MakingFairnessMedical Diagnosis+1Can I Trust My Fairness Metric? Assessing Fairness with Unlabeled Data and Bayesian Inference
We investigate the problem of reliably assessing group fairness when labeled examples are few but unlabeled examples are plentiful. We propose a general Bayesian framework that can augment labeled data with unlabeled dat…
Bayesian InferenceFairnessLeveraging Semi-Supervised Learning for Fairness using Neural Networks
There has been a growing concern about the fairness of decision-making systems based on machine learning. The shortage of labeled data has been always a challenging problem facing machine learning based systems. In such …
BIG-bench Machine LearningDecision MakingFairnessA Self-Supervised Learning Pipeline for Demographically Fair Facial Attribute Classification
Published research highlights the presence of demographic bias in automated facial attribute classification. The proposed bias mitigation techniques are mostly based on supervised learning, which requires a large amount …
AttributeContrastive LearningFacial Attribute ClassificationFairness+4Don't Throw it Away! The Utility of Unlabeled Data in Fair Decision Making
Decision making algorithms, in practice, are often trained on data that exhibits a variety of biases. Decision-makers often aim to take decisions based on some ground-truth target that is assumed or expected to be unbias…
Decision MakingFairness