paper-with-me

홈 › Papers

Concept-Based Explainable Artificial Intelligence: Metrics and Benchmarks

2025-01-31 · Halil Ibrahim Aysel, Xiaohao Cai, Adam Prugel-Bennett

Concept-based explanation methods, such as concept bottleneck models (CBMs), aim to improve the interpretability of machine learning models by linking their decisions to human-understandable concepts, under the critical assumption that such concepts can be accurately attributed to the network's feature space. However, this foundational assumption has not been rigorously validated, mainly because the field lacks standardised metrics and benchmarks to assess the existence and spatial alignment of such concepts. To address this, we propose three metrics: the concept global importance metric, the concept existence metric, and the concept location metric, including a technique for visualising concept activations, i.e., concept activation mapping. We benchmark post-hoc CBMs to illustrate their capabilities and challenges. Through qualitative and quantitative experiments, we demonstrate that, in many cases, even the most important concepts determined by post-hoc CBMs are not present in input images; moreover, when they are present, their saliency maps fail to align with the expected regions by either activating across an entire object or misidentifying relevant concept-specific regions. We analyse the root causes of these limitations, such as the natural correlation of concepts. Our findings underscore the need for more careful application of concept-based explanation techniques especially in settings where spatial interpretability is critical.

📄 PDF Abstract BibTeX arXiv:2501.19271

Code (0)

등록된 구현이 없습니다.

Tasks

Explainable artificial intelligence

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Concept-based Explainable Artificial Intelligence: A Survey

2023-12-20 · Eleonora Poeta, Gabriele Ciravegna, Eliana Pastor, Tania Cerquitelli 외

The field of explainable artificial intelligence emerged in response to the growing need for more transparent and reliable models. However, using raw features to provide explanations has been disputed in several works la…

Explainable artificial intelligenceSurvey

Comprehensible Artificial Intelligence on Knowledge Graphs: A survey

2024-04-04 · Simon Schramm, Christoph Wehner, Ute Schmid

Artificial Intelligence applications gradually move outside the safe walls of research labs and invade our daily lives. This is also true for Machine Learning methods on Knowledge Graphs, which has led to a steady increa…

Explainable artificial intelligenceInterpretable Machine LearningKnowledge GraphsSurvey

Explainable Artificial Intelligence for Assault Sentence Prediction in New Zealand

2022-08-15 · Harry Rodger, Andrew Lensen, Marcin Betkier

The judiciary has historically been conservative in its use of Artificial Intelligence, but recent advances in machine learning have prompted scholars to reconsider such use in tasks like sentence prediction. This paper …

Explainable artificial intelligenceSentence

Automated Molecular Concept Generation and Labeling with Large Language Models

2024-06-13 · Zimin Zhang, Qianli Wu, Botao Xia, Fang Sun 외

Artificial intelligence (AI) is transforming scientific research, with explainable AI methods like concept-based models (CMs) showing promise for new discoveries. However, in molecular science, CMs are less common than b…

In-Context Learning

A Review of Explainable Artificial Intelligence in Manufacturing

2021-07-05 · Georgios Sofianidis, Jože M. Rožanec, Dunja Mladenić, Dimosthenis Kyriazis

The implementation of Artificial Intelligence (AI) systems in the manufacturing domain enables higher production efficiency, outstanding performance, and safer operations, leveraging powerful tools such as deep learning …

Decision MakingExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)reinforcement-learning+1