Pre or Post-Softmax Scores in Gradient-based Attribution Methods, What is Best?
Gradient based attribution methods for neural networks working as classifiers use gradients of network scores. Here we discuss the practical differences between using gradients of pre-softmax scores versus post-softmax scores, and their respective advantages and disadvantages.
Code (1)
Similar Papers 제목 키워드 기반
Greedy PIG: Adaptive Integrated Gradients
Deep learning has become the standard approach for most machine learning tasks. While its impact is undeniable, interpreting the predictions of deep learning models from a human perspective remains a challenge. In contra…
Deep Learningfeature selectionA Polynomial Architecture-Attribution Co-Design Framework for Exact Aumann-Shapley Attribution in GNNs
We study feature-level and node-level explanations for graph neural networks (GNNs) through the lens of Aumann-Shapley attribution. Path-integral methods such as Integrated Gradients provide an axiomatic formulation of a…
A Vulnerability of Attribution Methods Using Pre-Softmax Scores
We discuss a vulnerability involving a category of attribution methods used to provide explanations for the outputs of convolutional neural networks working as classifiers. It is known that this type of networks are vuln…
Towards trustworthy explanations with gradient-based attribution methods
The low interpretability of deep neural networks (DNNs) remains a key barrier to their wide-spread adoption in the sciences. Attribution methods offer a promising solution, providing feature importance scores that serve …
Feature ImportanceModel Selectionscientific discoveryDynamic Top-k Estimation Consolidates Disagreement between Feature Attribution Methods
Feature attribution scores are used for explaining the prediction of a text classifier to users by highlighting a k number of tokens. In this work, we propose a way to determine the number of optimal k tokens that should…
Sentence