Mitigating belief projection in explainable artificial intelligence via Bayesian Teaching
State-of-the-art deep-learning systems use decision rules that are challenging for humans to model. Explainable AI (XAI) attempts to improve human understanding but rarely accounts for how people typically reason about unfamiliar agents. We propose explicitly modeling the human explainee via Bayesian Teaching, which evaluates explanations by how much they shift explainees' inferences toward a desired goal. We assess Bayesian Teaching in a binary image classification task across a variety of contexts. Absent intervention, participants predict that the AI's classifications will match their own, but explanations generated by Bayesian Teaching improve their ability to predict the AI's judgements by moving them away from this prior belief. Bayesian Teaching further allows each case to be broken down into sub-examples (here saliency maps). These sub-examples complement whole examples by improving error detection for familiar categories, whereas whole examples help predict correct AI judgements of unfamiliar cases.
Code (1)
Tasks
Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)image-classificationImage ClassificationSimilar Papers 제목 키워드 기반
Levels of explainable artificial intelligence for human-aligned conversational explanations
Over the last few years there has been rapid research growth into eXplainable Artificial Intelligence (XAI) and the closely aligned Interpretable Machine Learning (IML). Drivers for this growth include recent legislative…
Decision MakingExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Interpretable Machine LearningIsopignistic Canonical Decomposition via Belief Evolution Network
Developing a general information processing model in uncertain environments is fundamental for the advancement of explainable artificial intelligence. Dempster-Shafer theory of evidence is a well-known and effective reas…
Explainable artificial intelligenceModular Belief Updates and Confusion about Measures of Certainty in Artificial Intelligence Research
Over the last decade, there has been growing interest in the use or measures or change in belief for reasoning with uncertainty in artificial intelligence research. An important characteristic of several methodologies th…
An Explainable Artificial Intelligence Framework for Quality-Aware IoE Service Delivery
One of the core envisions of the sixth-generation (6G) wireless networks is to accumulate artificial intelligence (AI) for autonomous controlling of the Internet of Everything (IoE). Particularly, the quality of IoE serv…
Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)regressionExplainable Goal-Driven Agents and Robots -- A Comprehensive Review
Recent applications of autonomous agents and robots, such as self-driving cars, scenario-based trainers, exploration robots, and service robots have brought attention to crucial trust-related challenges associated with t…
Continual LearningExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Self-Driving Cars