A Formal Approach to Explainability
We regard explanations as a blending of the input sample and the model's output and offer a few definitions that capture various desired properties of the function that generates these explanations. We study the links between these properties and between explanation-generating functions and intermediate representations of learned models and are able to show, for example, that if the activations of a given layer are consistent with an explanation, then so do all other subsequent layers. In addition, we study the intersection and union of explanations as a way to construct new explanations.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
From Robustness to Explainability and Back Again
Formal explainability guarantees the rigor of computed explanations, and so it is paramount in domains where rigor is critical, including those deemed high-risk. Unfortunately, since its inception formal explainability h…
Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Machine Reasoning Explainability
As a field of AI, Machine Reasoning (MR) uses largely symbolic means to formalize and emulate abstract reasoning. Studies in early MR have notably started inquiries into Explainable AI (XAI) -- arguably one of the bigges…
Explainable Artificial Intelligence (XAI)Explaining Machine Learning Models using Entropic Variable Projection
In this paper, we present a new explainability formalism designed to shed light on how each input variable of a test set impacts the predictions of machine learning models. Hence, we propose a group explainability formal…
BIG-bench Machine LearningTowards Quantification of Explainability in Explainable Artificial Intelligence Methods
Artificial Intelligence (AI) has become an integral part of domains such as security, finance, healthcare, medicine, and criminal justice. Explaining the decisions of AI systems in human terms is a key challenge--due to …
Explainable artificial intelligenceUncovering Bugs in Formal Explainers: A Case Study with PyXAI
Formal explainable artificial intelligence (XAI) offers unique theoretical guarantees of rigor when compared to other non-formal methods of explainability. However, little attention has been given to the validation of pr…