Deep Fair Discriminative Clustering
Deep clustering has the potential to learn a strong representation and hence better clustering performance compared to traditional clustering methods such as $k$-means and spectral clustering. However, this strong representation learning ability may make the clustering unfair by discovering surrogates for protected information which we empirically show in our experiments. In this work, we study a general notion of group-level fairness for both binary and multi-state protected status variables (PSVs). We begin by formulating the group-level fairness problem as an integer linear programming formulation whose totally unimodular constraint matrix means it can be efficiently solved via linear programming. We then show how to inject this solver into a discriminative deep clustering backbone and hence propose a refinement learning algorithm to combine the clustering goal with the fairness objective to learn fair clusters adaptively. Experimental results on real-world datasets demonstrate that our model consistently outperforms state-of-the-art fair clustering algorithms. Our framework shows promising results for novel clustering tasks including flexible fairness constraints, multi-state PSVs and predictive clustering.
Code (1)
Tasks
ClusteringDeep ClusteringFairnessRepresentation LearningSimilar Papers 제목 키워드 기반
Towards Fair Deep Clustering With Multi-State Protected Variables
Fair clustering under the disparate impact doctrine requires that population of each protected group should be approximately equal in every cluster. Previous work investigated a difficult-to-scale pre-processing step for…
AttributeClusteringDeep ClusteringFairnessDeep Fair Multi-View Clustering with Attention KAN
Multi-view clustering is effective in unsupervised multi-view data analysis and has received considerable attention. However, most existing methods excessively emphasize certain attributes, resulting in unfair cluste…
ClusteringFairnessKolmogorov-Arnold NetworksDiscriminative Entropy Clustering and its Relation to K-means and SVM
Maximization of mutual information between the model's input and output is formally related to "decisiveness" and "fairness" of the softmax predictions, motivating these unsupervised entropy-based criteria for clustering…
ClusteringDeep ClusteringFairnessPseudo Label+1FACROC: a fairness measure for FAir Clustering through ROC curves
Fair clustering has attracted remarkable attention from the research community. Many fairness measures for clustering have been proposed; however, they do not take into account the clustering quality w.r.t. the values of…
AttributeClusteringFairnessRobust Fair Clustering: A Novel Fairness Attack and Defense Framework
Clustering algorithms are widely used in many societal resource allocation applications, such as loan approvals and candidate recruitment, among others, and hence, biased or unfair model outputs can adversely impact indi…
Adversarial AttackClusteringFairnessgraph partitioning