paper-with-me

홈 › Papers

Explainability's Gain is Optimality's Loss? -- How Explanations Bias Decision-making

2022-06-17 · Charles Wan, Rodrigo Belo, Leid Zejnilović

Decisions in organizations are about evaluating alternatives and choosing the one that would best serve organizational goals. To the extent that the evaluation of alternatives could be formulated as a predictive task with appropriate metrics, machine learning algorithms are increasingly being used to improve the efficiency of the process. Explanations help to facilitate communication between the algorithm and the human decision-maker, making it easier for the latter to interpret and make decisions on the basis of predictions by the former. Feature-based explanations' semantics of causal models, however, induce leakage from the decision-maker's prior beliefs. Our findings from a field experiment demonstrate empirically how this leads to confirmation bias and disparate impact on the decision-maker's confidence in the predictions. Such differences can lead to sub-optimal and biased decision outcomes.

📄 PDF Abstract BibTeX arXiv:2206.08705

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Fairness and Explainability: Bridging the Gap Towards Fair Model Explanations

2022-12-07 · Yuying Zhao, Yu Wang, Tyler Derr

While machine learning models have achieved unprecedented success in real-world applications, they might make biased/unfair decisions for specific demographic groups and hence result in discriminative outcomes. Although …

Decision MakingFairness

Gender Bias in Explainability: Investigating Performance Disparity in Post-hoc Methods

2025-05-02 · Mahdi Dhaini, Ege Erdogan, Nils Feldhus, Gjergji Kasneci

While research on applications and evaluations of explanation methods continues to expand, fairness of the explanation methods concerning disparities in their performance across subgroups remains an often overlooked aspe…

Fairness

Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?

2025-09-26 · Yifan Wang, Mayank Jobanputra, Ji-Ung Lee, Soyoung Oh 외 arxiv

Natural language processing (NLP) models often replicate or amplify social bias from training data, raising concerns about fairness. At the same time, their black-box nature makes it difficult for users to recognize bias…

Hate Speech Detection

Beyond Trivial Counterfactual Explanations with Diverse Valuable Explanations

2021-03-18 · ICCV 2021 10 · Pau Rodriguez, Massimo Caccia, Alexandre Lacoste, Lee Zamparo 외

Explainability for machine learning models has gained considerable attention within the research community given the importance of deploying more reliable machine-learning systems. In computer vision applications, genera…

AttributeBIG-bench Machine LearningcounterfactualDecision Making+1

Constructing Fair Latent Space for Intersection of Fairness and Explainability

2024-12-23 · Hyungjun Joo, Hyeonggeun Han, Sehwan Kim, Sangwoo Hong 외

As the use of machine learning models has increased, numerous studies have aimed to enhance fairness. However, research on the intersection of fairness and explainability remains insufficient, leading to potential issues…

counterfactualFairness