paper-with-me

Papers

FACE: Faithful Automatic Concept Extraction

2025-10-13 · Dipkamal Bhusal, Michael Clifford, Sara Rampazzi, Nidhi Rastogi arxiv

Interpreting deep neural networks through concept-based explanations offers a bridge between low-level features and high-level human-understandable semantics. However, existing automatic concept discovery methods often fail to align these extracted concepts with the model's true decision-making process, thereby compromising explanation faithfulness. In this work, we propose FACE (Faithful Automatic Concept Extraction), a novel framework that augments Non-negative Matrix Factorization (NMF) with a Kullback-Leibler (KL) divergence regularization term to ensure alignment between the model's original and concept-based predictions. Unlike prior methods that operate solely on encoder activations, FACE incorporates classifier supervision during concept learning, enforcing predictive consistency and enabling faithful explanations. We provide theoretical guarantees showing that minimizing the KL divergence bounds the deviation in predictive distributions, thereby promoting faithful local linearity in the learned concept space. Systematic evaluations on ImageNet, COCO, and CelebA datasets demonstrate that FACE outperforms existing methods across faithfulness and sparsity metrics.

📄 PDF Abstract BibTeX arXiv:2510.11675

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CRAFT: Concept Recursive Activation FacTorization for Explainability

2022-11-17 · CVPR 2023 1 · Thomas Fel, Agustin Picard, Louis Bethune, Thibaut Boissin 외

Attribution methods, which employ heatmaps to identify the most influential regions of an image that impact model decisions, have gained widespread popularity as a type of explainability method. However, recent research …

EviSnap: Faithful Evidence-Cited Explanations for Cold-Start Cross-Domain Recommendation

2026-01-09 · Yingjun Dai, Ahmed El-Roby arxiv

Cold-start cross-domain recommender (CDR) systems predict a user's preferences in a target domain using only their source-domain behavior, yet existing CDR models either map opaque embeddings or rely on post-hoc or LLM-g…

Mapping Faithful Reasoning in Language Models

2025-10-25 · Jiazheng Li, Andreas Damianou, J Rosser, José Luis Redondo García 외 arxiv

Chain-of-thought (CoT) traces promise transparency for reasoning language models, but prior work shows they are not always faithful reflections of internal computation. This raises challenges for oversight: practitioners…

High-Precision Extraction of Emerging Concepts from Scientific Literature

2020-06-11 · Daniel King, Doug Downey, Daniel S. Weld

Identification of new concepts in scientific literature can help power faceted search, scientific trend analysis, knowledge-base construction, and more, but current methods are lacking. Manual identification cannot keep …

Knowledge Base ConstructionVocal Bursts Intensity Prediction

Naming the Concepts Classifiers Rely On: Language-Anchored Decomposition for Faithful Explanation

2026-07-08 · Ahsan Habib Akash, Dipkamal Bhusal, Stacey Jones, Donald A. Adjeroh 외 arxiv

Deep neural networks are widely deployed in high-stakes visual applications where interpretability is critical, yet existing explanations face a trade-off: post-hoc concept methods recover factors that are faithful to a …