paper-with-me

홈 › Papers

Post-hoc Self-explanation of CNNs

2026-03-30 · Ahcène Boubekki, Line H. Clemmensen arxiv

Although standard Convolutional Neural Networks (CNNs) can be mathematically reinterpreted as Self-Explainable Models (SEMs), their built-in prototypes do not on their own accurately represent the data. Replacing the final linear layer with a $k$-means-based classifier addresses this limitation without compromising performance. This work introduces a common formalization of $k$-means-based post-hoc explanations for the classifier, the encoder's final output (B4), and combinations of intermediate feature activations. The latter approach leverages the spatial consistency of convolutional receptive fields to generate concept-based explanation maps, which are supported by gradient-free feature attribution maps. Empirical evaluation with a ResNet34 shows that using shallower, less compressed feature activations, such as those from the last three blocks (B234), results in a trade-off between semantic fidelity and a slight reduction in predictive performance.

📄 PDF Abstract BibTeX arXiv:2603.28466

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TTE-CAM: Self-Explainable Class Activation Maps for Pretrained Black-Box CNNs

2026-03-27 · Kerol Djoumessi, Philipp Berens arxiv

Convolutional neural networks (CNNs) achieve state-of-the-art performance in medical image analysis yet remain opaque, limiting adoption in high-stakes clinical settings. Existing approaches face a fundamental trade-off:…

Learning Robust Convolutional Neural Networks with Relevant Feature Focusing via Explanations

2022-02-09 · Kazuki Adachi, Shin'ya Yamaguchi

Existing image recognition techniques based on convolutional neural networks (CNNs) basically assume that the training and test datasets are sampled from i.i.d distributions. However, this assumption is easily broken in …

PACE: Posthoc Architecture-Agnostic Concept Extractor for Explaining CNNs

2021-08-31 · Vidhya Kamakshi, Uday Gupta, Narayanan C Krishnan

Deep CNNs, though have achieved the state of the art performance in image classification tasks, remain a black-box to a human using them. There is a growing interest in explaining the working of these deep models to impr…

image-classificationImage Classification

Why Self-Attention? A Targeted Evaluation of Neural Machine Translation Architectures

2018-08-27 · EMNLP 2018 10 · Gongbo Tang, Mathias Müller, Annette Rios, Rico Sennrich

Recently, non-recurrent architectures (convolutional, self-attentional) have outperformed RNNs in neural machine translation. CNNs and self-attentional networks can connect distant words via shorter network paths than RN…

Machine TranslationTranslationWord Sense Disambiguation

Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations

2024-02-19 · Milan Bhan, Jean-Noel Vittaut, Nicolas Chesneau, Marie-Jeanne Lesot

Incorporating natural language rationales in the prompt and In-Context Learning (ICL) have led to a significant improvement of Large Language Models (LLMs) performance. However, generating high-quality rationales require…

In-Context Learning