paper-with-me

Papers

Reason induced visual attention for explainable autonomous driving

2021-10-11 · Sikai Chen, Jiqian Dong, Runjia Du, Yujie Li, Samuel Labi

Deep learning (DL) based computer vision (CV) models are generally considered as black boxes due to poor interpretability. This limitation impedes efficient diagnoses or predictions of system failure, thereby precluding the widespread deployment of DLCV models in safety-critical tasks such as autonomous driving. This study is motivated by the need to enhance the interpretability of DL model in autonomous driving and therefore proposes an explainable DL-based framework that generates textual descriptions of the driving environment and makes appropriate decisions based on the generated descriptions. The proposed framework imitates the learning process of human drivers by jointly modeling the visual input (images) and natural language, while using the language to induce the visual attention in the image. The results indicate strong explainability of autonomous driving decisions obtained by focusing on relevant features from visual inputs. Furthermore, the output attention maps enhance the interpretability of the model not only by providing meaningful explanation to the model behavior but also by identifying the weakness of and potential improvement directions for the model.

📄 PDF Abstract BibTeX arXiv:2110.07380

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

Development and testing of an image transformer for explainable autonomous driving systems

2021-10-11 · Jiqian Dong, Sikai Chen, Shuya Zong, Tiantian Chen 외

In the last decade, deep learning (DL) approaches have been used successfully in computer vision (CV) applications. However, DL-based CV models are generally considered to be black boxes due to their lack of interpretabi…

Autonomous DrivingDiagnostic

Where, What, Why: Towards Explainable Driver Attention Prediction

2025-06-29 · Yuchen Zhou, Jiayu Tang, Xiaoyan Xiao, Yueyao Lin 외

Modeling task-driven attention in driving is a fundamental challenge for both autonomous vehicles and cognitive science. Existing methods primarily predict where drivers look by generating spatial heatmaps, but fail to c…

Autonomous DrivingAutonomous VehiclesDriver Attention MonitoringLarge Language Model+2

Interpretable Visual Reasoning via Induced Symbolic Space

2020-11-23 · ICCV 2021 10 · Zhonghao Wang, Kai Wang, Mo Yu, JinJun Xiong 외

We study the problem of concept induction in visual reasoning, i.e., identifying concepts and their hierarchical relationships from question-answer pairs associated with images; and achieve an interpretable model via wor…

Visual Question Answering (VQA)Visual Reasoning

Explainable Multi-Camera 3D Object Detection with Transformer-Based Saliency Maps

2023-12-22 · Till Beemelmanns, Wassim Zahr, Lutz Eckstein

Vision Transformers (ViTs) have achieved state-of-the-art results on various computer vision tasks, including 3D object detection. However, their end-to-end implementation also makes ViTs less explainable, which can be a…

3D Object DetectionAutonomous Drivingobject-detectionObject Detection

Explainable and Explicit Visual Reasoning over Scene Graphs

2018-12-05 · CVPR 2019 6 · Jiaxin Shi, Hanwang Zhang, Juanzi Li

We aim to dismantle the prevalent black-box neural architectures used in complex visual reasoning tasks, into the proposed eXplainable and eXplicit Neural Modules (XNMs), which advance beyond existing neural module netwo…

Inductive BiasVisual Question Answering (VQA)Visual Reasoning