paper-with-me

Papers

The Disagreement Problem in Explainable Machine Learning: A Practitioner's Perspective

2022-02-03 · Satyapriya Krishna, Tessa Han, Alex Gu, Steven Wu, Shahin Jabbari, Himabindu Lakkaraju

As various post hoc explanation methods are increasingly being leveraged to explain complex models in high-stakes settings, it becomes critical to develop a deeper understanding of whether and when the explanations output by these methods disagree with each other, and how such disagreements are resolved in practice. However, there is little to no research that provides answers to these critical questions. In this work, we formalize and study the disagreement problem in explainable machine learning. More specifically, we define the notion of disagreement between explanations, analyze how often such disagreements occur in practice, and how practitioners resolve these disagreements. We first conduct interviews with data scientists to understand what constitutes disagreement between explanations generated by different methods for the same model prediction, and introduce a novel quantitative framework to formalize this understanding. We then leverage this framework to carry out a rigorous empirical analysis with four real-world datasets, six state-of-the-art post hoc explanation methods, and six different predictive models, to measure the extent of disagreement between the explanations generated by various popular explanation methods. In addition, we carry out an online user study with data scientists to understand how they resolve the aforementioned disagreements. Our results indicate that (1) state-of-the-art explanation methods often disagree in terms of the explanations they output, and (2) machine learning practitioners often employ ad hoc heuristics when resolving such disagreements. These findings suggest that practitioners may be relying on misleading explanations when making consequential decisions. They also underscore the importance of developing principled frameworks for effectively evaluating and comparing explanations output by various explanation techniques.

📄 PDF Abstract BibTeX arXiv:2202.01602

Code (1)

grobruegge/vitexplcomp pytorch

Tasks

BIG-bench Machine Learning

Methods 이 논문이 사용한 방법론

HOC 설명 없음

Similar Papers 제목 키워드 기반

EXAGREE: Towards Explanation Agreement in Explainable Machine Learning

2024-11-04 · Sichao Li, Quanling Deng, Amanda S. Barnard

Explanations in machine learning are critical for trust, transparency, and fairness. Yet, complex disagreements among these explanations limit the reliability and applicability of machine learning models, especially in h…

Fairness

The Meta-Evaluation Problem in Explainable AI: Identifying Reliable Estimators with MetaQuantus

2023-02-14 · Anna Hedström, Philine Bommer, Kristoffer K. Wickstrøm, Wojciech Samek 외

One of the unsolved challenges in the field of Explainable AI (XAI) is determining how to most reliably estimate the quality of an explanation method in the absence of ground truth explanation labels. Resolving this issu…

Explainable Artificial Intelligence (XAI)

Explainable News Summarization -- Analysis and mitigation of Disagreement Problem

2024-10-24 · Seema Aswani, Sujala D. Shetty

Explainable AI (XAI) techniques for text summarization provide valuable understanding of how the summaries are generated. Recent studies have highlighted a major challenge in this area, known as the disagreement problem.…

Extreme SummarizationNews SummarizationText Summarization

Adoption of Explainable Natural Language Processing: Perspectives from Industry and Academia on Practices and Challenges

2025-08-13 · Mahdi Dhaini, Tobias Müller, Roksoliana Rabets, Gjergji Kasneci arxiv

The field of explainable natural language processing (NLP) has grown rapidly in recent years. The growing opacity of complex models calls for transparency and explanations of their decisions, which is crucial to understa…

Manipulation Risks in Explainable AI: The Implications of the Disagreement Problem

2023-06-24 · Sofie Goethals, David Martens, Theodoros Evgeniou

Artificial Intelligence (AI) systems are increasingly used in high-stakes domains of our life, increasing the need to explain these decisions and to make sure that they are aligned with how we want the decision to be mad…

Explainable Artificial Intelligence (XAI)