Interpretability Benchmark for Evaluating Spatial Misalignment of Prototypical Parts Explanations
Prototypical parts-based networks are becoming increasingly popular due to their faithful self-explanations. However, their similarity maps are calculated in the penultimate network layer. Therefore, the receptive field of the prototype activation region often depends on parts of the image outside this region, which can lead to misleading interpretations. We name this undesired behavior a spatial explanation misalignment and introduce an interpretability benchmark with a set of dedicated metrics for quantifying this phenomenon. In addition, we propose a method for misalignment compensation and apply it to existing state-of-the-art models. We show the expressiveness of our benchmark and the effectiveness of the proposed compensation methodology through extensive empirical studies.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Interpretable Few-Shot Image Classification via Prototypical Concept-Guided Mixture of LoRA Experts
Self-Explainable Models (SEMs) rely on Prototypical Concept Learning (PCL) to enable their visual recognition processes more interpretable, but they often struggle in data-scarce settings where insufficient training samp…
Explainable ModelsFew-Shot Image Classificationimage-classificationImage ClassificationExpert-Guided Explainable Few-Shot Learning with Active Sample Selection for Medical Image Analysis
Medical image analysis faces two critical challenges: scarcity of labeled data and lack of model interpretability, both hindering clinical AI deployment. Few-shot learning (FSL) addresses data limitations but lacks trans…
Few-Shot LearningActive LearningProtoQuant: Quantization of Prototypical Parts For General and Fine-Grained Image Classification
Prototypical parts-based models offer a "this looks like that" paradigm for intrinsic interpretability, yet they typically struggle with ImageNet-scale generalization and often require computationally expensive backbone …
Fine-Grained Image ClassificationDeformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes
We present a deformable prototypical part network (Deformable ProtoPNet), an interpretable image classifier that integrates the power of deep learning and the interpretability of case-based reasoning. This model classifi…
FairnessImage ClassificationDeep Unfolding Network with Spatial Alignment for multi-modal MRI reconstruction
Multi-modal Magnetic Resonance Imaging (MRI) offers complementary diagnostic information, but some modalities are limited by the long scanning time. To accelerate the whole acquisition process, MRI reconstruction of one …
DiagnosticMRI Reconstruction