paper-with-me

홈 › Papers

Evaluating the Stability of Semantic Concept Representations in CNNs for Robust Explainability

2023-04-28 · Georgii Mikriukov, Gesina Schwalbe, Christian Hellert, Korinna Bade

Analysis of how semantic concepts are represented within Convolutional Neural Networks (CNNs) is a widely used approach in Explainable Artificial Intelligence (XAI) for interpreting CNNs. A motivation is the need for transparency in safety-critical AI-based systems, as mandated in various domains like automated driving. However, to use the concept representations for safety-relevant purposes, like inspection or error retrieval, these must be of high quality and, in particular, stable. This paper focuses on two stability goals when working with concept representations in computer vision CNNs: stability of concept retrieval and of concept attribution. The guiding use-case is a post-hoc explainability framework for object detection (OD) CNNs, towards which existing concept analysis (CA) methods are successfully adapted. To address concept retrieval stability, we propose a novel metric that considers both concept separation and consistency, and is agnostic to layer and concept representation dimensionality. We then investigate impacts of concept abstraction level, number of concept training samples, CNN size, and concept representation dimensionality on stability. For concept attribution stability we explore the effect of gradient instability on gradient-based explainability methods. The results on various CNNs for classification and object detection yield the main findings that (1) the stability of concept retrieval can be enhanced through dimensionality reduction via data aggregation, and (2) in shallow layers where gradient instability is more pronounced, gradient smoothing techniques are advised. Finally, our approach provides valuable insights into selecting the appropriate layer and concept representation dimensionality, paving the way towards CA in safety-critical XAI applications.

📄 PDF Abstract BibTeX arXiv:2304.14864

Code (0)

등록된 구현이 없습니다.

Tasks

Dimensionality ReductionExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)object-detectionObject DetectionRetrieval

Similar Papers 제목 키워드 기반

Unsupervised learning of object semantic parts from internal states of CNNs by population encoding

2015-11-21 · Jianyu Wang, Zhishuai Zhang, Cihang Xie, Vittal Premachandran 외

We address the key question of how object part representations can be found from the internal states of CNNs that are trained for high-level tasks, such as object classification. This work provides a new unsupervised met…

ClusteringKeypoint DetectionObject

Network Dissection: Quantifying Interpretability of Deep Visual Representations

2017-04-19 · CVPR 2017 7 · David Bau, Bolei Zhou, Aditya Khosla, Aude Oliva 외

We propose a general framework called Network Dissection for quantifying the interpretability of latent representations of CNNs by evaluating the alignment between individual hidden units and a set of semantic concepts. …

Revealing Similar Semantics Inside CNNs: An Interpretable Concept-based Comparison of Feature Spaces

2023-04-30 · Georgii Mikriukov, Gesina Schwalbe, Christian Hellert, Korinna Bade

Safety-critical applications require transparency in artificial intelligence (AI) components, but widely used convolutional neural networks (CNNs) widely used for perception tasks lack inherent interpretability. Hence, i…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Model Selection

ConceptFlow: Hierarchical and Fine-grained Concept-Based Explanation for Convolutional Neural Networks

2025-09-16 · Xinyu Mu, Hui Dou, Furao Shen, Jian Zhao arxiv

Concept-based interpretability for Convolutional Neural Networks (CNNs) aims to align internal model representations with high-level semantic concepts, but existing approaches largely overlook the semantic roles of indiv…

Evaluating distributed word representations for capturing semantics of biomedical concepts

2015-07-01 · WS 2015 7 · MUNEEB TH, Sunil Sahu, Ashish Anand
ArticlesChunkingLanguage ModelingLanguage Modelling+6