Contrastive Explanations for Explaining Model Adaptations
Many decision making systems deployed in the real world are not static - a phenomenon known as model adaptation takes place over time. The need for transparency and interpretability of AI-based decision models is widely accepted and thus have been worked on extensively. Usually, explanation methods assume a static system that has to be explained. Explaining non-static systems is still an open research question, which poses the challenge how to explain model adaptations. In this contribution, we propose and (empirically) evaluate a framework for explaining model adaptations by contrastive explanations. We also propose a method for automatically finding regions in data space that are affected by a given model adaptation and thus should be explained.
Code (1)
Tasks
Decision MakingmodelSimilar Papers 제목 키워드 기반
Explaining NLP Models via Minimal Contrastive Editing (MiCE)
Humans have been shown to give contrastive explanations, which explain why an observed event happened rather than some other counterfactual event (the contrast case). Despite the influential role that contrastivity plays…
counterfactualMultiple-choiceQuestion AnsweringSentiment Analysis+2Towards Transparent Robotic Planning via Contrastive Explanations
Providing explanations of chosen robotic actions can help to increase the transparency of robotic planning and improve users' trust. Social sciences suggest that the best explanations are contrastive, explaining not just…
(When) Are Contrastive Explanations of Reinforcement Learning Helpful?
Global explanations of a reinforcement learning (RL) agent's expected behavior can make it safer to deploy. However, such explanations are often difficult to understand because of the complicated nature of many RL polici…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Model-contrastive explanations through symbolic reasoning
Explaining how two machine learning classification models differ in their behaviour is gaining significance in eXplainable AI, given the increasing diffusion of learning-based decision support systems. Human decision-mak…
Counterfactual ExplanationExplainable artificial intelligencemodelC-SENN: Contrastive Self-Explaining Neural Network
In this study, we use a self-explaining neural network (SENN), which learns unsupervised concepts, to acquire concepts that are easy for people to understand automatically. In concept learning, the hidden layer retains v…
Autonomous DrivingContrastive Learning