Towards Fully Interpretable Deep Neural Networks: Are We There Yet?
Despite the remarkable performance, Deep Neural Networks (DNNs) behave as black-boxes hindering user trust in Artificial Intelligence (AI) systems. Research on opening black-box DNN can be broadly categorized into post-hoc methods and inherently interpretable DNNs. While many surveys have been conducted on post-hoc interpretation methods, little effort is devoted to inherently interpretable DNNs. This paper provides a review of existing methods to develop DNNs with intrinsic interpretability, with a focus on Convolutional Neural Networks (CNNs). The aim is to understand the current progress towards fully interpretable DNNs that can cater to different interpretation requirements. Finally, we identify gaps in current work and suggest potential research directions.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Residualized Similarity for Faithfully Explainable Authorship Verification
Responsible use of Authorship Verification (AV) systems not only requires high accuracy but also interpretable solutions. More importantly, for systems to be used to make decisions with real-world consequences requires t…
Structural Neural Additive Models: Enhanced Interpretable Machine Learning
Deep neural networks (DNNs) have shown exceptional performances in a wide range of tasks and have become the go-to method for problems requiring high-level predictive power. There has been extensive research on how DNNs …
Additive modelsInterpretable Machine LearningFocusLearn: Fully-Interpretable, High-Performance Modular Neural Networks for Time Series
Multivariate time series have many applications, from healthcare and meteorology to life science. Although deep learning models have shown excellent predictive performance for time series, they have been criticised for b…
Additive modelsfeature selectionTime SeriesTime Series Forecasting+1A Hybrid Fully Convolutional CNN-Transformer Model for Inherently Interpretable Medical Image Classification
In many medical imaging tasks, convolutional neural networks (CNNs) efficiently extract local features hierarchically. More recently, vision transformers (ViTs) have gained popularity, using self-attention mechanisms to …
image-classificationImage ClassificationMedical Image ClassificationVariational pSOM: Deep Probabilistic Clustering with Self-Organizing Maps
Generating visualizations and interpretations from high-dimensional data is a common problem in many fields. Two key approaches for tackling this problem are clustering and representation learning. There are very perfor…
ClusteringDeep ClusteringRepresentation LearningTime Series+1