paper-with-me

홈 › Papers

Cause and Effect: Hierarchical Concept-based Explanation of Neural Networks

2021-05-14 · Mohammad Nokhbeh Zaeem, Majid Komeili

In many scenarios, human decisions are explained based on some high-level concepts. In this work, we take a step in the interpretability of neural networks by examining their internal representation or neuron's activations against concepts. A concept is characterized by a set of samples that have specific features in common. We propose a framework to check the existence of a causal relationship between a concept (or its negation) and task classes. While the previous methods focus on the importance of a concept to a task class, we go further and introduce four measures to quantitatively determine the order of causality. Moreover, we propose a method for constructing a hierarchy of concepts in the form of a concept-based decision tree which can shed light on how various concepts interact inside a neural network towards predicting output classes. Through experiments, we demonstrate the effectiveness of the proposed method in explaining the causal relationship between a concept and the predictive behaviour of a neural network as well as determining the interactions between different concepts through constructing a concept hierarchy.

📄 PDF Abstract BibTeX arXiv:2105.07033

Code (0)

등록된 구현이 없습니다.

Tasks

Negation

Similar Papers 제목 키워드 기반

Hierarchical Concept Embedding & Pursuit for Interpretable Image Classification

2026-02-11 · Nghia Nguyen, Tianjiao Ding, René Vidal arxiv

Interpretable-by-design models are gaining traction in computer vision because they provide faithful explanations for their predictions. In image classification, these models typically recover human-interpretable concept…

Image Classification

Hierarchical, Interpretable, Label-Free Concept Bottleneck Model

2026-04-02 · Haodong Xie, Yujun Cai, Rahul Singh Maharjan, Yiwei Wang 외 arxiv

Concept Bottleneck Models (CBMs) introduce interpretability to black-box deep learning models by predicting labels through human-understandable concepts. However, unlike humans, who identify objects at different levels o…

ConceptFlow: Hierarchical and Fine-grained Concept-Based Explanation for Convolutional Neural Networks

2025-09-16 · Xinyu Mu, Hui Dou, Furao Shen, Jian Zhao arxiv

Concept-based interpretability for Convolutional Neural Networks (CNNs) aims to align internal model representations with high-level semantic concepts, but existing approaches largely overlook the semantic roles of indiv…

Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations

2025-04-19 · Katie Matton, Robert Osazuwa Ness, John Guttag, Emre Kiciman

Large language models (LLMs) are capable of generating plausible explanations of how they arrived at an answer to a question. However, these explanations can misrepresent the model's "reasoning" process, i.e., they can b…

Language ModelingLanguage ModellingLarge Language ModelMedical Question Answering+1

Hierarchical Attention Network for Explainable Depression Detection on Twitter Aided by Metaphor Concept Mappings

2022-09-15 · COLING 2022 10 · Sooji Han, Rui Mao, Erik Cambria

Automatic depression detection on Twitter can help individuals privately and conveniently understand their mental health status in the early stages before seeing mental health professionals. Most existing black-box-like …

Decision MakingDepression Detection