paper-with-me

Papers

Marrying Fairness and Explainability in Supervised Learning

2022-04-06 · Przemyslaw Grabowicz, Nicholas Perello, Aarshee Mishra

Machine learning algorithms that aid human decision-making may inadvertently discriminate against certain protected groups. We formalize direct discrimination as a direct causal effect of the protected attributes on the decisions, while induced discrimination as a change in the causal influence of non-protected features associated with the protected attributes. The measurements of marginal direct effect (MDE) and SHapley Additive exPlanations (SHAP) reveal that state-of-the-art fair learning methods can induce discrimination via association or reverse discrimination in synthetic and real-world datasets. To inhibit discrimination in algorithmic systems, we propose to nullify the influence of the protected attribute on the output of the system, while preserving the influence of remaining features. We introduce and study post-processing methods achieving such objectives, finding that they yield relatively high model accuracy, prevent direct discrimination, and diminishes various disparity measures, e.g., demographic disparity.

📄 PDF Abstract BibTeX arXiv:2204.02947

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeDecision MakingFairness

Similar Papers 제목 키워드 기반

Explainability for fair machine learning

2020-10-14 · Tom Begley, Tobias Schwedes, Christopher Frye, Ilya Feige

As the decisions made or influenced by machine learning models increasingly impact our lives, it is crucial to detect, understand, and mitigate unfairness. But even simply determining what "unfairness" should mean in a g…

AttributeBIG-bench Machine LearningFairness

Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?

2025-09-26 · Yifan Wang, Mayank Jobanputra, Ji-Ung Lee, Soyoung Oh 외 arxiv

Natural language processing (NLP) models often replicate or amplify social bias from training data, raising concerns about fairness. At the same time, their black-box nature makes it difficult for users to recognize bias…

Hate Speech Detection

Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models

2022-06-08 · NAACL (TrustNLP) 2022 7 · Esma Balkir, Svetlana Kiritchenko, Isar Nejadgholi, Kathleen C. Fraser

Motivations for methods in explainable artificial intelligence (XAI) often include detecting, quantifying and mitigating bias, and contributing to making machine learning models fairer. However, exactly how an XAI method…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Fairness

Fairness and Explainability in Automatic Decision-Making Systems. A challenge for computer science and law

2022-05-14 · Thierry Kirst, Olivia Tambou, Virginie Do, Alexis Tsoukiàs

The paper offers a contribution to the interdisciplinary constructs of analyzing fairness issues in automatic algorithmic decisions. Section 1 shows that technical choices in supervised learning have social implications …

Decision MakingFairness

Harnessing value from data science in business: ensuring explainability and fairness of solutions

2021-08-10 · Krzysztof Chomiak, Michał Miktus

The paper introduces concepts of fairness and explainability (XAI) in artificial intelligence, oriented to solve a sophisticated business problems. For fairness, the authors discuss the bias-inducing specifics, as well a…

Explainable Artificial Intelligence (XAI)Fairness