paper-with-me

홈 › Papers

Natural Images are More Informative for Interpreting CNN Activations than State-of-the-Art Synthetic Feature Visualizations

2020-10-09 · NeurIPS Workshop SVRHM 2020 12 · Judy Borowski, Roland Simon Zimmermann, Judith Schepers, Robert Geirhos, Thomas S. A. Wallis, Matthias Bethge, Wieland Brendel

Feature visualizations such as synthetic maximally activating images are a widely used explanation method to better understand the information processing of convo- lutional neural networks (CNNs). At the same time, there are concerns that these visualizations might not accurately represent CNNs’ inner workings. Here, we measure how much extremely activating images help humans in predicting CNN activations. Using a well-controlled psychophysical paradigm, we compare the informativeness of synthetic images by Olah et al. [45] with a simple baseline visualization, namely natural images that also strongly activate a specific feature map. Given either synthetic or natural reference images, human participants choose which of two query images leads to strong positive activation. The experiment is designed to maximize participants’ performance, and is the first to probe interme- diate instead of final layer representations. We find that synthetic images indeed provide helpful information about feature map activations (82 ± 4% accuracy; chance would be 50%). However, natural images—originally intended to be a baseline—outperform these synthetic images by a wide margin (92 ± 2% accuracy). The superiority of natural images holds across the investigated network and various conditions. Therefore, we argue that visualization methods should improve over this simple baseline.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Informativeness

Similar Papers 제목 키워드 기반

Exemplary natural images explain CNN activations better than synthetic feature visualizations

2021-01-01 · ICLR 2021 1 · Judy Borowski, Roland Simon Zimmermann, Judith Schepers, Robert Geirhos 외

Feature visualizations such as synthetic maximally activating images are a widely used explanation method to better understand the information processing of convolutional neural networks (CNNs). At the same time, there a…

Informativeness

Exemplary Natural Images Explain CNN Activations Better than State-of-the-Art Feature Visualization

2020-10-23 · Judy Borowski, Roland S. Zimmermann, Judith Schepers, Robert Geirhos 외

Feature visualizations such as synthetic maximally activating images are a widely used explanation method to better understand the information processing of convolutional neural networks (CNNs). At the same time, there a…

Informativeness

Understanding Neural Networks Through Deep Visualization

2015-06-22 · Jason Yosinski, Jeff Clune, Anh Nguyen, Thomas Fuchs 외

Recent years have produced great advances in training large, deep neural networks (DNNs), including notable successes in training convolutional neural networks (convnets) to recognize natural images. However, our underst…

Interpretable Machine Learning

Interpretability as Compression: Reconsidering SAE Explanations of Neural Activations with MDL-SAEs

2024-10-15 · Kola Ayonrinde, Michael T. Pearce, Lee Sharkey

Sparse Autoencoders (SAEs) have emerged as a useful tool for interpreting the internal representations of neural networks. However, naively optimising SAEs for reconstruction loss and sparsity results in a preference for…

Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants

2025-12-17 · Vincent Huang, Dami Choi, Daniel D. Johnson, Sarah Schwettmann 외 arxiv

Interpreting the internal activations of neural networks can produce more faithful explanations of their behavior, but is difficult due to the complex structure of activation space. Existing approaches to scalable interp…