Do Not Trust Additive Explanations
Explainable Artificial Intelligence (XAI)has received a great deal of attention recently. Explainability is being presented as a remedy for the distrust of complex and opaque models. Model agnostic methods such as LIME, SHAP, or Break Down promise instance-level interpretability for any complex machine learning model. But how faithful are these additive explanations? Can we rely on additive explanations for non-additive models? In this paper, we (1) examine the behavior of the most popular instance-level explanations under the presence of interactions, (2) introduce a new method that detects interactions for instance-level explanations, (3) perform a large scale benchmark to see how frequently additive explanations may be misleading.
Code (2)
Tasks
Additive modelsBIG-bench Machine LearningExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Considerations When Learning Additive Explanations for Black-Box Models
Many methods to explain black-box models, whether local or global, are additive. In this paper, we study global additive explanations for non-additive models, focusing on four explanation methods: partial dependence, Sha…
Additive modelsAn evaluation of quality and robustness of smoothed explanations
Explanation methods play a crucial role in helping to understand the decisions of deep neural networks (DNNs) to develop trust that is critical for the adoption of predictive models. However, explanation methods are easi…
SHAP-Based Explanation Methods: A Review for NLP Interpretability
Model explanations are crucial for the transparent, safe, and trustworthy deployment of machine learning models. The SHapley Additive exPlanations (SHAP) framework is considered by many to be a gold standard for local ex…
SHAP-Based Explanation Methods: A Review for NLP Interpretability
Model explanations are crucial for the transparent, safe, and trustworthy deployment of machine learning models. The \emph{SHapley Additive exPlanations} (SHAP) framework is considered by many to be a gold standard for l…
Explainable deepfake and spoofing detection: an attack analysis using SHapley Additive exPlanations
Despite several years of research in deepfake and spoofing detection for automatic speaker verification, little is known about the artefacts that classifiers use to distinguish between bona fide and spoofed utterances. A…
Face SwappingSpeaker Verification