paper-with-me

홈 › Papers

Better Understanding Differences in Attribution Methods via Systematic Evaluations

2023-03-21 · Sukrut Rao, Moritz Böhle, Bernt Schiele

Deep neural networks are very successful on many vision tasks, but hard to interpret due to their black box nature. To overcome this, various post-hoc attribution methods have been proposed to identify image regions most influential to the models' decisions. Evaluating such methods is challenging since no ground truth attributions exist. We thus propose three novel evaluation schemes to more reliably measure the faithfulness of those methods, to make comparisons between them more fair, and to make visual inspection more systematic. To address faithfulness, we propose a novel evaluation setting (DiFull) in which we carefully control which parts of the input can influence the output in order to distinguish possible from impossible attributions. To address fairness, we note that different methods are applied at different layers, which skews any comparison, and so evaluate all methods on the same layers (ML-Att) and discuss how this impacts their performance on quantitative metrics. For more systematic visualizations, we propose a scheme (AggAtt) to qualitatively evaluate the methods on complete datasets. We use these evaluation schemes to study strengths and shortcomings of some widely used attribution methods over a wide range of models. Finally, we propose a post-processing smoothing step that significantly improves the performance of some attribution methods, and discuss its applicability.

📄 PDF Abstract BibTeX arXiv:2303.11884

Code (1)

sukrutrao/attribution-evaluation 공식 구현 pytorch

Tasks

Fairness

Similar Papers 제목 키워드 기반

Understanding Structured Health Data through Interaction-Aware Mixture-of-Experts

2026-07-14 · Ji Hwan Park, Ying Ding, Tianjin Guo arxiv

We study interaction-aware mixture-of-experts for post-stroke rigidity prediction using multi-level views of structured health records. Despite minimal performance gains, routing attribution reveals systematic importance…

Discriminative Attribution from Counterfactuals

2021-09-28 · Nils Eckstein, Alexander S. Bates, Gregory S. X. E. Jefferis, Jan Funke

We present a method for neural network interpretability by combining feature attribution with counterfactual explanations to generate attribution maps that highlight the most discriminative features between pairs of clas…

counterfactual

Attribution Explanations for Deep Neural Networks: A Theoretical Perspective

2025-08-11 · Huiqi Deng, Hongbin Pei, Quanshi Zhang, Mengnan Du arxiv

Attribution explanation is a typical approach for explaining deep neural networks (DNNs), inferring an importance or contribution score for each input variable to the final output. In recent years, numerous attribution m…

Towards Better Understanding Attribution Methods

2022-05-20 · CVPR 2022 1 · Sukrut Rao, Moritz Böhle, Bernt Schiele

Deep neural networks are very successful on many vision tasks, but hard to interpret due to their black box nature. To overcome this, various post-hoc attribution methods have been proposed to identify image regions most…

Explainable artificial intelligenceExplanation Fidelity EvaluationImage ClassificationInterpretable Machine Learning

IMACS: Image Model Attribution Comparison Summaries

2022-01-26 · Eldon Schoop, Ben Wedin, Andrei Kapishnikov, Tolga Bolukbasi 외

Developing a suitable Deep Neural Network (DNN) often requires significant iteration, where different model versions are evaluated and compared. While metrics such as accuracy are a powerful means to succinctly describe …

image-classificationImage Classificationmodel