paper-with-me

Papers

How can we trust opaque systems? Criteria for robust explanations in XAI

2025-08-18 · Florian J. Boge, Annika Schuster arxiv

Deep learning (DL) algorithms are becoming ubiquitous in everyday life and in scientific research. However, the price we pay for their impressively accurate predictions is significant: their inner workings are notoriously opaque - it is unknown to laypeople and researchers alike what features of the data a DL system focuses on and how it ultimately succeeds in predicting correct outputs. A necessary criterion for trustworthy explanations is that they should reflect the relevant processes the algorithms' predictions are based on. The field of eXplainable Artificial Intelligence (XAI) presents promising methods to create such explanations. But recent reviews about their performance offer reasons for skepticism. As we will argue, a good criterion for trustworthiness is explanatory robustness: different XAI methods produce the same explanations in comparable contexts. However, in some instances, all methods may give the same, but still wrong, explanation. We therefore argue that in addition to explanatory robustness (ER), a prior requirement of explanation method robustness (EMR) has to be fulfilled by every XAI method. Conversely, the robustness of an individual method is in itself insufficient for trustworthiness. In what follows, we develop and formalize criteria for ER as well as EMR, providing a framework for explaining and establishing trust in DL algorithms. We also highlight interesting application cases and outline directions for future work.

📄 PDF Abstract BibTeX arXiv:2508.12623

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Meta Survey of Quality Evaluation Criteria in Explanation Methods

2022-03-25 · Helena Löfström, Karl Hammar, Ulf Johansson

Explanation methods and their evaluation have become a significant issue in explainable artificial intelligence (XAI) due to the recent surge of opaque AI models in decision support systems (DSS). Since the most accurate…

Bias DetectionExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Survey

CHILLI: A data context-aware perturbation method for XAI

2024-07-10 · Saif Anwar, Nathan Griffiths, Abhir Bhalerao, Thomas Popham

The trustworthiness of Machine Learning (ML) models can be difficult to assess, but is critical in high-risk or ethically sensitive applications. Many models are treated as a `black-box' where the reasoning or criteria f…

Explaining Explanations to Society

2019-01-19 · Leilani H. Gilpin, Cecilia Testart, Nathaniel Fruchter, Julius Adebayo

There is a disconnect between explanatory artificial intelligence (XAI) methods and the types of explanations that are useful for and demanded by society (policy makers, government officials, etc.) Questions that experts…

Decision MakingExplainable Artificial Intelligence (XAI)

LIMEADE: From AI Explanations to Advice Taking

2020-03-09 · Benjamin Charles Germain Lee, Doug Downey, Kyle Lo, Daniel S. Weld

Research in human-centered AI has shown the benefits of systems that can explain their predictions. Methods that allow an AI to take advice from humans in response to explanations are similarly useful. While both capabil…

BIG-bench Machine Learningimage-classificationImage ClassificationRecommendation Systems

Black Box Deployed -- Functional Criteria for Artificial Moral Agents in the LLM Era

2025-07-17 · Matthew E. Brophy arxiv

The advancement of powerful yet opaque large language models (LLMs) necessitates a fundamental revision of the philosophical criteria used to evaluate artificial moral agents (AMAs). Pre-LLM frameworks often relied on th…