paper-with-me

홈 › Papers

Selective Explanations: Leveraging Human Input to Align Explainable AI

2023-01-23 · Vivian Lai, Yiming Zhang, Chacha Chen, Q. Vera Liao, Chenhao Tan

While a vast collection of explainable AI (XAI) algorithms have been developed in recent years, they are often criticized for significant gaps with how humans produce and consume explanations. As a result, current XAI techniques are often found to be hard to use and lack effectiveness. In this work, we attempt to close these gaps by making AI explanations selective -- a fundamental property of human explanations -- by selectively presenting a subset from a large set of model reasons based on what aligns with the recipient's preferences. We propose a general framework for generating selective explanations by leveraging human input on a small sample. This framework opens up a rich design space that accounts for different selectivity goals, types of input, and more. As a showcase, we use a decision-support task to explore selective explanations based on what the decision-maker would consider relevant to the decision task. We conducted two experimental studies to examine three out of a broader possible set of paradigms based on our proposed framework: in Study 1, we ask the participants to provide their own input to generate selective explanations, with either open-ended or critique-based input. In Study 2, we show participants selective explanations based on input from a panel of similar users (annotators). Our experiments demonstrate the promise of selective explanations in reducing over-reliance on AI and improving decision outcomes and subjective perceptions of the AI, but also paint a nuanced picture that attributes some of these positive effects to the opportunity to provide one's own input to augment AI explanations. Overall, our work proposes a novel XAI framework inspired by human communication behaviors and demonstrates its potentials to encourage future work to better align AI explanations with human production and consumption of explanations.

📄 PDF Abstract BibTeX arXiv:2301.09656

Code (0)

등록된 구현이 없습니다.

Tasks

Explainable Artificial Intelligence (XAI)

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Talking Back -- human input and explanations to interactive AI systems

2025-03-06 · Alan Dix, Tommaso Turchi, Ben Wilson, Anna Monreale 외

While XAI focuses on providing AI explanations to humans, can the reverse - humans explaining their judgments to AI - foster richer, synergistic human-AI systems? This paper explores various forms of human inputs to AI a…

Robust Counterfactual Explanations on Graph Neural Networks

2021-07-08 · NeurIPS 2021 12 · Mohit Bajaj, Lingyang Chu, Zi Yu Xue, Jian Pei 외

Massive deployment of Graph Neural Networks (GNNs) in high-stake applications generates a strong demand for explanations that are robust to noise and align well with human intuition. Most existing methods generate explan…

counterfactualPrediction

Structure Your Data: Towards Semantic Graph Counterfactuals

2024-03-11 · Angeliki Dimitriou, Maria Lymperaiou, Giorgos Filandrianos, Konstantinos Thomas 외

Counterfactual explanations (CEs) based on concepts are explanations that consider alternative scenarios to understand which high-level semantic features contributed to particular model predictions. In this work, we prop…

counterfactualDescriptiveGraph Similarity

Explaining Motion Relevance for Activity Recognition in Video Deep Learning Models

2020-03-31 · Liam Hiley, Alun Preece, Yulia Hicks, Supriyo Chakraborty 외

A small subset of explainability techniques developed initially for image recognition models has recently been applied for interpretability of 3D Convolutional Neural Network models in activity recognition tasks. Much li…

Activity Recognition

To what extent do human explanations of model behavior align with actual model behavior?

2020-12-24 · EMNLP (BlackboxNLP) 2021 11 · Grusha Prasad, Yixin Nie, Mohit Bansal, Robin Jia 외

Given the increasingly prominent role NLP models (will) play in our lives, it is important for human expectations of model behavior to align with actual model behavior. Using Natural Language Inference (NLI) as a case st…

modelNatural Language Inference