paper-with-me

홈 › Papers

Explaining Explanations: An Overview of Interpretability of Machine Learning

2018-05-31 · Leilani H. Gilpin, David Bau, Ben Z. Yuan, Ayesha Bajwa, Michael Specter, Lalana Kagal

There has recently been a surge of work in explanatory artificial intelligence (XAI). This research area tackles the important problem that complex machines and algorithms often cannot provide insights into their behavior and thought processes. XAI allows users and parts of the internal system to be more transparent, providing explanations of their decisions in some level of detail. These explanations are important to ensure algorithmic fairness, identify potential bias/problems in the training data, and to ensure that the algorithms perform as expected. However, explanations produced by these systems is neither standardized nor systematically assessed. In an effort to create best practices and identify open challenges, we provide our definition of explainability and show how it can be used to classify existing literature. We discuss why current approaches to explanatory methods especially for deep neural networks are insufficient. Finally, based on our survey, we conclude with suggested future research directions for explanatory artificial intelligence.

📄 PDF Abstract BibTeX arXiv:1806.00069

Code (1)

Mahdidrm/Emotion-Recognition tf

Tasks

BIG-bench Machine LearningExplainable Artificial Intelligence (XAI)Fairness

Similar Papers 제목 키워드 기반

Towards Robust Interpretability with Self-Explaining Neural Networks

2018-12-01 · NeurIPS 2018 12 · David Alvarez Melis, Tommi Jaakkola

Most recent work on interpretability of complex machine learning models has focused on estimating a-posteriori explanations for previously trained models around specific predictions. Self-explaining models where interpre…

Towards Robust Interpretability with Self-Explaining Neural Networks

2018-06-20 · NeurIPS 2018 · David Alvarez-Melis, Tommi S. Jaakkola

Most recent work on interpretability of complex machine learning models has focused on estimating $\textit{a posteriori}$ explanations for previously trained models around specific predictions. $\textit{Self-explaining}$…

Explaining Deep Neural Networks and Beyond: A Review of Methods and Applications

2020-03-17 · Wojciech Samek, Grégoire Montavon, Sebastian Lapuschkin, Christopher J. Anders 외

With the broader and highly successful usage of machine learning in industry and the sciences, there has been a growing demand for Explainable AI. Interpretability and explanation methods for gaining a better understandi…

BIG-bench Machine LearningInterpretable Machine Learning

Learning by Self-Explaining

2023-09-15 · Wolfgang Stammer, Felix Friedrich, David Steinmann, Manuel Brack 외

Much of explainable AI research treats explanations as a means for model inspection. Yet, this neglects findings from human psychology that describe the benefit of self-explanations in an agent's learning process. Motiva…

image-classificationImage Classification

Analyzing the Interpretability Robustness of Self-Explaining Models

2019-05-27 · Haizhong Zheng, Earlence Fernandes, Atul Prakash

Recently, interpretable models called self-explaining models (SEMs) have been proposed with the goal of providing interpretability robustness. We evaluate the interpretability robustness of SEMs and show that explanation…