paper-with-me

Papers

Fidelity of Interpretability Methods and Perturbation Artifacts in Neural Networks

2022-03-06 · Lennart Brocki, Neo Christopher Chung

Despite excellent performance of deep neural networks (DNNs) in image classification, detection, and prediction, characterizing how DNNs make a given decision remains an open problem, resulting in a number of interpretability methods. Post-hoc interpretability methods primarily aim to quantify the importance of input features with respect to the class probabilities. However, due to the lack of ground truth and the existence of interpretability methods with diverse operating characteristics, evaluating these methods is a crucial challenge. A popular approach to evaluate interpretability methods is to perturb input features deemed important for a given prediction and observe the decrease in accuracy. However, perturbation itself may introduce artifacts. We propose a method for estimating the impact of such artifacts on the fidelity estimation by utilizing model accuracy curves from perturbing input features according to the Most Import First (MIF) and Least Import First (LIF) orders. Using the ResNet-50 trained on the ImageNet, we demonstrate the proposed fidelity estimation of four popular post-hoc interpretability methods.

📄 PDF Abstract BibTeX arXiv:2203.02928

Code (0)

등록된 구현이 없습니다.

Tasks

image-classificationImage Classification

Similar Papers 제목 키워드 기반

Feature Perturbation Augmentation for Reliable Evaluation of Importance Estimators in Neural Networks

2023-03-02 · Lennart Brocki, Neo Christopher Chung

Post-hoc explanation methods attempt to make the inner workings of deep neural networks more interpretable. However, since a ground truth is in general lacking, local post-hoc interpretability methods, which assign impor…

Data AugmentationPrediction

Counterfactual Generation with Knockoffs

2021-02-01 · Oana-Iuliana Popescu, Maha Shadaydeh, Joachim Denzler

Human interpretability of deep neural networks' decisions is crucial, especially in domains where these directly affect human lives. Counterfactual explanations of already trained neural networks can be generated by pert…

counterfactualVariable Selection

Interpretable and Robust AI in EEG Systems: A Survey

2023-04-21 · Xinliang Zhou, Chenyu Liu, Zhongruo Wang, Liming Zhai 외

The close coupling of artificial intelligence (AI) and electroencephalography (EEG) has substantially advanced human-computer interaction (HCI) technologies in the AI era. Different from traditional EEG systems, the inte…

EEGSurvey

Sparsity-Driven Parallel Imaging Consistency for Improved Self-Supervised MRI Reconstruction

2025-05-30 · Yaşar Utku Alçalar, Mehmet Akçakaya

Physics-driven deep learning (PD-DL) models have proven to be a powerful approach for improved reconstruction of rapid MRI scans. In order to train these models in scenarios where fully-sampled reference data is unavaila…

MRI ReconstructionSelf-Supervised Learning

XAI-CLIP: ROI-Guided Perturbation Framework for Explainable Medical Image Segmentation in Multimodal Vision-Language Models

2026-02-01 · Thuraya Alzubaidi, Sana Ammar, Maryam Alsharqi, Islem Rekik 외 arxiv

Medical image segmentation is a critical component of clinical workflows, enabling accurate diagnosis, treatment planning, and disease monitoring. However, despite the superior performance of transformer-based models ove…

Medical Image Segmentation