paper-with-me

홈 › Papers

Towards Better Understanding Attribution Methods

2022-05-20 · CVPR 2022 1 · Sukrut Rao, Moritz Böhle, Bernt Schiele

Deep neural networks are very successful on many vision tasks, but hard to interpret due to their black box nature. To overcome this, various post-hoc attribution methods have been proposed to identify image regions most influential to the models' decisions. Evaluating such methods is challenging since no ground truth attributions exist. We thus propose three novel evaluation schemes to more reliably measure the faithfulness of those methods, to make comparisons between them more fair, and to make visual inspection more systematic. To address faithfulness, we propose a novel evaluation setting (DiFull) in which we carefully control which parts of the input can influence the output in order to distinguish possible from impossible attributions. To address fairness, we note that different methods are applied at different layers, which skews any comparison, and so evaluate all methods on the same layers (ML-Att) and discuss how this impacts their performance on quantitative metrics. For more systematic visualizations, we propose a scheme (AggAtt) to qualitatively evaluate the methods on complete datasets. We use these evaluation schemes to study strengths and shortcomings of some widely used attribution methods. Finally, we propose a post-processing smoothing step that significantly improves the performance of some attribution methods, and discuss its applicability.

📄 PDF Abstract BibTeX arXiv:2205.10435

Code (1)

sukrutrao/attribution-evaluation 공식 구현 pytorch

Tasks

Explainable artificial intelligenceExplanation Fidelity EvaluationImage ClassificationInterpretable Machine Learning

Similar Papers 제목 키워드 기반

XRAI: Better Attributions Through Regions

2019-06-06 · ICCV 2019 10 · Andrei Kapishnikov, Tolga Bolukbasi, Fernanda Viégas, Michael Terry

Saliency methods can aid understanding of deep neural networks. Recent years have witnessed many improvements to saliency methods, as well as new ways for evaluating them. In this paper, we 1) present a novel region-base…

Towards better understanding of gradient-based attribution methods for Deep Neural Networks

2017-11-16 · ICLR 2018 1 · Marco Ancona, Enea Ceolini, Cengiz Öztireli, Markus Gross

Understanding the flow of information in Deep Neural Networks (DNNs) is a challenging problem that has gain increasing attention over the last few years. While several methods have been proposed to explain network predic…

text-classificationText Classification

Discriminative Attribution from Counterfactuals

2021-09-28 · Nils Eckstein, Alexander S. Bates, Gregory S. X. E. Jefferis, Jan Funke

We present a method for neural network interpretability by combining feature attribution with counterfactual explanations to generate attribution maps that highlight the most discriminative features between pairs of clas…

counterfactual

Time-series attribution maps with regularized contrastive learning

2025-02-17 · Steffen Schneider, Rodrigo González Laiz, Anastasiia Filippova, Markus Frey 외

Gradient-based attribution methods aim to explain decisions of deep learning models but so far lack identifiability guarantees. Here, we propose a method to generate attribution maps with identifiability guarantees by de…

Contrastive LearningTime Series

Towards a better understanding of Burrows's Delta in literary authorship attribution

2015-06-01 · WS 2015 6 · Stefan Evert, Thomas Proisl, Thorsten Vitt, Christof Sch{\"o}ch 외
Authorship AttributionText Clustering