A Rate-Distortion Framework for Explaining Black-box Model Decisions
We present the Rate-Distortion Explanation (RDE) framework, a mathematically well-founded method for explaining black-box model decisions. The framework is based on perturbations of the target input signal and applies to any differentiable pre-trained model such as neural networks. Our experiments demonstrate the framework's adaptability to diverse data modalities, particularly images, audio, and physical simulations of urban environments.
Code (0)
등록된 구현이 없습니다.
Tasks
Physical SimulationsSimilar Papers 제목 키워드 기반
Stop Explaining Black Box Machine Learning Models for High Stakes Decisions and Use Interpretable Models Instead
Black box machine learning models are currently being used for high stakes decision-making throughout society, causing problems throughout healthcare, criminal justice, and in other domains. People have hoped that creati…
BIG-bench Machine LearningDecision MakingInterpretable Machine LearningExplainers in the Wild: Making Surrogate Explainers Robust to Distortions through Perception
Explaining the decisions of models is becoming pervasive in the image processing domain, whether it is by using post-hoc methods or by creating inherently interpretable models. While the widespread use of surrogate expla…
image-classificationImage ClassificationA Rate-Distortion Framework for Explaining Neural Network Decisions
We formalise the widespread idea of interpreting neural network decisions as an explicit optimisation problem in a rate-distortion framework. A set of input features is deemed relevant for a classification decision if th…
General Classificationimage-classificationImage ClassificationTimeSHAP: Explaining Recurrent Models through Sequence Perturbations
Recurrent neural networks are a standard building block in numerous machine learning domains, from natural language processing to time-series classification. While their application has grown ubiquitous, understanding of…
Decision MakingFeature ImportanceTime SeriesTime Series Analysis+1Gradient Frequency Modulation for Visually Explaining Video Understanding Models
In many applications, it is essential to understand why a machine learning model makes the decisions it does, but this is inhibited by the black-box nature of state-of-the-art neural networks. Because of this, increasing…
Action RecognitionTemporal Action LocalizationVideo Understanding