Individual Fairness Revisited: Transferring Techniques from Adversarial Robustness
We turn the definition of individual fairness on its head---rather than ascertaining the fairness of a model given a predetermined metric, we find a metric for a given model that satisfies individual fairness. This can facilitate the discussion on the fairness of a model, addressing the issue that it may be difficult to specify a priori a suitable metric. Our contributions are twofold: First, we introduce the definition of a minimal metric and characterize the behavior of models in terms of minimal metrics. Second, for more complicated models, we apply the mechanism of randomized smoothing from adversarial robustness to make them individually fair under a given weighted $L^p$ metric. Our experiments show that adapting the minimal metrics of linear models to more complicated neural networks can lead to meaningful and interpretable fairness guarantees at little cost to utility.
Code (0)
등록된 구현이 없습니다.
Tasks
Adversarial RobustnessFairnessMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
A Tale of Fairness Revisited: Beyond Adversarial Learning for Deep Neural Network Fairness
Motivated by the need for fair algorithmic decision making in the age of automation and artificially-intelligent technology, this technical report provides a theoretical insight into adversarial training for fairness in …
Decision MakingFairnessBetter Algorithms for Individually Fair $k$-Clustering
We study data clustering problems with $\ell_p$-norm objectives (e.g. $k$-Median and $k$-Means) in the context of individual fairness. The dataset consists of $n$ points, and we want to find $k$ centers such that (a) the…
ClusteringFairnessIndividual Fairness in Bayesian Neural Networks
We study Individual Fairness (IF) for Bayesian neural networks (BNNs). Specifically, we consider the $\epsilon$-$\delta$-individual fairness notion, which requires that, for any pair of input points that are $\epsilon$-s…
Adversarial RobustnessBayesian InferenceFairnessThe Double-Edged Sword of Input Perturbations to Robust Accurate Fairness
Deep neural networks (DNNs) are known to be sensitive to adversarial input perturbations, leading to a reduction in either prediction accuracy or individual fairness. To jointly characterize the susceptibility of predict…
Adversarial AttackFairnessCausal Adversarial Perturbations for Individual Fairness and Robustness in Heterogeneous Data Spaces
As responsible AI gains importance in machine learning algorithms, properties such as fairness, adversarial robustness, and causality have received considerable attention in recent years. However, despite their individua…
Adversarial RobustnessFairnessSemantic SimilaritySemantic Textual Similarity