paper-with-me

Papers

UFO: A unified method for controlling Understandability and Faithfulness Objectives in concept-based explanations for CNNs

2023-03-27 · Vikram V. Ramaswamy, Sunnie S. Y. Kim, Ruth Fong, Olga Russakovsky

Concept-based explanations for convolutional neural networks (CNNs) aim to explain model behavior and outputs using a pre-defined set of semantic concepts (e.g., the model recognizes scene class `bedroom'' based on the presence of concepts bed'' and pillow''). However, they often do not faithfully (i.e., accurately) characterize the model's behavior and can be too complex for people to understand. Further, little is known about how faithful and understandable different explanation methods are, and how to control these two properties. In this work, we propose UFO, a unified method for controlling Understandability and Faithfulness Objectives in concept-based explanations. UFO formalizes understandability and faithfulness as mathematical objectives and unifies most existing concept-based explanations methods for CNNs. Using UFO, we systematically investigate how explanations change as we turn the knobs of faithfulness and understandability. Our experiments demonstrate a faithfulness-vs-understandability tradeoff: increasing understandability reduces faithfulness. We also provide insights into the `disagreement problem'' in explainable machine learning, by analyzing when and how concept-based explanations disagree with each other.

📄 PDF Abstract BibTeX arXiv:2303.15632

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

LCE: A Framework for Explainability of DNNs for Ultrasound Image Based on Concept Discovery

2024-08-19 · Weiji Kong, Xun Gong, Juan Wang

Explaining the decisions of Deep Neural Networks (DNNs) for medical images has become increasingly important. Existing attribution methods have difficulty explaining the meaning of pixels while existing concept-based met…

Diagnostic

Towards Human-Understandable Multi-Dimensional Concept Discovery

2025-03-24 · CVPR 2025 1 · Arne Grobrügge, Niklas Kühl, Gerhard Satzger, Philipp Spitzer

Concept-based eXplainable AI (C-XAI) aims to overcome the limitations of traditional saliency maps by converting pixels into human-understandable concepts that are consistent across an entire dataset. A crucial aspect of…

RFEval: Benchmarking Reasoning Faithfulness under Counterfactual Reasoning Intervention in Large Reasoning Models

2026-02-19 · Yunseok Han, Yejoon Lee, Jaeyoung Do arxiv

Large Reasoning Models (LRMs) exhibit strong performance, yet often produce rationales that sound plausible but fail to reflect their true decision process, undermining reliability and trust. We introduce a formal framew…

Evaluating Readability and Faithfulness of Concept-based Explanations

2024-04-29 · Meng Li, Haoran Jin, Ruixuan Huang, Zhihao Xu 외

With the growing popularity of general-purpose Large Language Models (LLMs), comes a need for more global explanations of model behaviors. Concept-based explanations arise as a promising avenue for explaining high-level …

A Causal Lens for Evaluating Faithfulness Metrics

2025-02-26 · Kerem Zaman, Shashank Srivastava

Large Language Models (LLMs) offer natural language explanations as an alternative to feature attribution methods for model interpretability. However, despite their plausibility, they may not reflect the model's internal…

Decision MakingFact CheckingModel EditingObject Counting