paper-with-me

홈 › Papers

Understanding Interpretability by generalized distillation in Supervised Classification

2020-12-05 · Adit Agarwal, Dr. K. K. Shukla, Arjan Kuijper, Anirban Mukhopadhyay

The ability to interpret decisions taken by Machine Learning (ML) models is fundamental to encourage trust and reliability in different practical applications. Recent interpretation strategies focus on human understanding of the underlying decision mechanisms of the complex ML models. However, these strategies are restricted by the subjective biases of humans. To dissociate from such human biases, we propose an interpretation-by-distillation formulation that is defined relative to other ML models. We generalize the distillation technique for quantifying interpretability, using an information-theoretic perspective, removing the role of ground-truth from the definition of interpretability. Our work defines the entropy of supervised classification models, providing bounds on the entropy of Piece-Wise Linear Neural Networks (PWLNs), along with the first theoretical bounds on the interpretability of PWLNs. We evaluate our proposed framework on the MNIST, Fashion-MNIST and Stanford40 datasets and demonstrate the applicability of the proposed theoretical framework in different supervised classification scenarios.

📄 PDF Abstract BibTeX arXiv:2012.03089

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classification

Methods 이 논문이 사용한 방법론

Interpretability 설명 없음

Similar Papers 제목 키워드 기반

Unifying distillation and privileged information

2015-11-11 · David Lopez-Paz, Léon Bottou, Bernhard Schölkopf, Vladimir Vapnik

Distillation (Hinton et al., 2015) and privileged information (Vapnik & Izmailov, 2015) are two techniques that enable machines to learn from other machines. This paper unifies these two techniques into generalized disti…

Class Attention Transfer Based Knowledge Distillation

2023-04-25 · CVPR 2023 1 · Ziyao Guo, Haonan Yan, Hui Li, Xiaodong Lin

Previous knowledge distillation methods have shown their impressive performance on model compression tasks, however, it is hard to explain how the knowledge they transferred helps to improve the performance of the studen…

Knowledge DistillationModel Compression

Injecting Explainability and Lightweight Design into Weakly Supervised Video Anomaly Detection Systems

2024-12-28 · Wen-Dong Jiang, Chih-Yung Chang, Hsiang-Chuan Chang, Ji-Yuan Chen 외

Weakly Supervised Monitoring Anomaly Detection (WSMAD) utilizes weak supervision learning to identify anomalies, a critical task for smart city monitoring. However, existing multimodal approaches often fail to meet the r…

Anomaly DetectionBinary ClassificationClassificationContrastive Learning+6

Interpretability-by-Design with Accurate Locally Additive Models and Conditional Feature Effects

2026-02-18 · Vasilis Gkolemis, Loukas Kavouras, Dimitrios Kyriakopoulos, Konstantinos Tsopelas 외 arxiv

Generalized additive models (GAMs) offer interpretability through independent univariate feature effects but underfit when interactions are present in data. GA$^2$Ms add selected pairwise interactions which improves accu…

Generalized Uncertainty of Deep Neural Networks: Taxonomy and Applications

2023-02-02 · chengyu dong

Deep neural networks have seen enormous success in various real-world applications. Beyond their predictions as point estimates, increasing attention has been focused on quantifying the uncertainty of their predictions. …

Knowledge DistillationModel CompressionUncertainty QuantificationWeakly-supervised Learning