paper-with-me

홈 › Papers

Multimodal Explanations: Justifying Decisions and Pointing to the Evidence

2018-02-15 · CVPR 2018 6 · Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata, Anna Rohrbach, Bernt Schiele, Trevor Darrell, Marcus Rohrbach

Deep models that are both effective and explainable are desirable in many settings; prior explainable models have been unimodal, offering either image-based visualization of attention weights or text-based generation of post-hoc justifications. We propose a multimodal approach to explanation, and argue that the two modalities provide complementary explanatory strengths. We collect two new datasets to define and evaluate this task, and propose a novel model which can provide joint textual rationale generation and attention visualization. Our datasets define visual and textual justifications of a classification decision for activity recognition tasks (ACT-X) and for visual question answering tasks (VQA-X). We quantitatively show that training with the textual explanations not only yields better textual justification models, but also better localizes the evidence that supports the decision. We also qualitatively show cases where visual explanation is more insightful than textual explanation, and vice versa, supporting our thesis that multimodal explanation models offer significant benefits over unimodal approaches.

📄 PDF Abstract BibTeX arXiv:1802.08129

Code (1)

Seth-Park/MultimodalExplanations 공식 구현 caffe2

Tasks

Activity RecognitionExplainable ModelsQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

Attentive Explanations: Justifying Decisions and Pointing to the Evidence (Extended Abstract)

2017-11-17 · Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata, Anna Rohrbach 외

Deep models are the defacto standard in visual decision problems due to their impressive performance on a wide array of visual tasks. On the other hand, their opaqueness has led to a surge of interest in explainable syst…

Question AnsweringVisual Question Answering (VQA)

Attentive Explanations: Justifying Decisions and Pointing to the Evidence

2016-12-14 · Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata, Bernt Schiele 외

Deep models are the defacto standard in visual decision models due to their impressive performance on a wide array of visual tasks. However, they are frequently seen as opaque and are unable to explain their decisions. I…

Decision MakingQuestion AnsweringSentenceVisual Question Answering+1

Where and When: Space-Time Attention for Audio-Visual Explanations

2021-05-04 · Yanbei Chen, Thomas Hummel, A. Sophia Koepke, Zeynep Akata

Explaining the decision of a multi-modal decision-maker requires to determine the evidence from both modalities. Recent advances in XAI provide explanations for models trained on still images. However, when it comes to m…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)Multimodal Deep Learning

Beyond Verification: Abductive Explanations for Post-AI Assessment of Privacy Leakage

2025-11-13 · Belona Sonna, Alban Grastien, Claire Benn arxiv

Privacy leakage in AI-based decision processes poses significant risks, particularly when sensitive information can be inferred. We propose a formal framework to audit privacy leakage using abductive explanations, which …

Explainable Decision Making with Lean and Argumentative Explanations

2022-01-18 · Xiuyi Fan, Francesca Toni

It is widely acknowledged that transparency of automated decision making is crucial for deployability of intelligent systems, and explaining the reasons why some decisions are "good" and some are not is a way to achievin…

Decision Making