paper-with-me

홈 › Papers

Local Explanation Methods for Deep Neural Networks Lack Sensitivity to Parameter Values

2018-10-08 · Julius Adebayo, Justin Gilmer, Ian Goodfellow, Been Kim

Explaining the output of a complicated machine learning model like a deep neural network (DNN) is a central challenge in machine learning. Several proposed local explanation methods address this issue by identifying what dimensions of a single input are most responsible for a DNN's output. The goal of this work is to assess the sensitivity of local explanations to DNN parameter values. Somewhat surprisingly, we find that DNNs with randomly-initialized weights produce explanations that are both visually and quantitatively similar to those produced by DNNs with learned weights. Our conjecture is that this phenomenon occurs because these explanations are dominated by the lower level features of a DNN, and that a DNN's architecture provides a strong prior which significantly affects the representations learned at these lower layers. NOTE: This work is now subsumed by our recent manuscript, Sanity Checks for Saliency Maps (to appear NIPS 2018), where we expand on findings and address concerns raised in Sundararajan et. al. (2018).

📄 PDF Abstract BibTeX arXiv:1810.03307

Code (2)

andresbecker/master_thesis tf
pytorch/captum pytorch

Tasks

BIG-bench Machine LearningSensitivity

Similar Papers 제목 키워드 기반

A Note about: Local Explanation Methods for Deep Neural Networks lack Sensitivity to Parameter Values

2018-06-11 · Mukund Sundararajan, Ankur Taly

Local explanation methods, also known as attribution methods, attribute a deep network's prediction to its input (cf. Baehrens et al. (2010)). We respond to the claim from Adebayo et al. (2018) that local explanation met…

AttributeSensitivity

Robustness of Explainable Artificial Intelligence in Industrial Process Modelling

2024-07-12 · Benedikt Kantz, Clemens Staudinger, Christoph Feilmayr, Johannes Wachlmayr 외

eXplainable Artificial Intelligence (XAI) aims at providing understandable explanations of black box models. In this paper, we evaluate current XAI methods by scoring them based on ground truth simulations and sensitivit…

ARCExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Sensitivity

Evaluating the overall sensitivity of saliency-based explanation methods

2023-06-21 · Harshinee Sriram, Cristina Conati

We address the need to generate faithful explanations of "black box" Deep Learning models. Several tests have been proposed to determine aspects of faithfulness of explanation methods, but they lack cross-domain applicab…

Sensitivity

SAM: The Sensitivity of Attribution Methods to Hyperparameters

2020-03-04 · CVPR 2020 6 · Naman Bansal, Chirag Agarwal, Anh Nguyen

Attribution methods can provide powerful insights into the reasons for a classifier's decision. We argue that a key desideratum of an explanation method is its robustness to input hyperparameters which are often randomly…

Sensitivity

On the (In)fidelity and Sensitivity of Explanations

2019-12-01 · NeurIPS 2019 12 · Chih-Kuan Yeh, Cheng-Yu Hsieh, Arun Suggala, David I. Inouye 외

We consider objective evaluation measures of saliency explanations for complex black-box machine learning models. We propose simple robust variants of two notions that have been considered in recent literature: (in)fidel…

Sensitivity