paper-with-me

홈 › Papers

Unifying Formal Explanations: A Complexity-Theoretic Perspective

2026-02-20 · Shahaf Bassan, Xuanxiang Huang, Guy Katz arxiv

Previous work has explored the computational complexity of deriving two fundamental types of explanations for ML model predictions: (1) *sufficient reasons*, which are subsets of input features that, when fixed, determine a prediction, and (2) *contrastive reasons*, which are subsets of input features that, when modified, alter a prediction. Prior studies have examined these explanations in different contexts, such as non-probabilistic versus probabilistic frameworks and local versus global settings. In this study, we introduce a unified framework for analyzing these explanations, demonstrating that they can all be characterized through the minimization of a unified probabilistic value function. We then prove that the complexity of these computations is influenced by three key properties of the value function: (1) *monotonicity*, (2) *submodularity*, and (3) *supermodularity* - which are three fundamental properties in *combinatorial optimization*. Our findings uncover some counterintuitive results regarding the nature of these properties within the explanation settings examined. For instance, although the *local* value functions do not exhibit monotonicity or submodularity/supermodularity whatsoever, we demonstrate that the *global* value functions do possess these properties. This distinction enables us to prove a series of novel polynomial-time results for computing various explanations with provable guarantees in the global explainability setting, across a range of ML models that span the interpretability spectrum, such as neural networks, decision trees, and tree ensembles. In contrast, we show that even highly simplified versions of these explanations become NP-hard to compute in the corresponding local explainability setting.

📄 PDF Abstract BibTeX arXiv:2602.18160

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations

2025-08-14 · Amir Mehrpanah, Matteo Gamba, Kevin Smith, Hossein Azizpour arxiv

ReLU networks, while prevalent for visual data, have sharp transitions, sometimes relying on individual pixels for predictions, making vanilla gradient-based explanations noisy and difficult to interpret. Existing method…

Local Explanations via Necessity and Sufficiency: Unifying Theory and Practice

2021-03-27 · David Watson, Limor Gultchin, Ankur Taly, Luciano Floridi

Necessity and sufficiency are the building blocks of all successful explanations. Yet despite their importance, these notions have been conceptually underdeveloped and inconsistently applied in explainable artificial int…

Explainable artificial intelligenceExplainable Artificial Intelligence (XAI)

Supervised Learning as Lossy Compression: Characterizing Generalization and Sample Complexity via Finite Blocklength Analysis

2026-02-04 · Kosuke Sugiyama, Masato Uchida arxiv

This paper presents a novel information-theoretic perspective on generalization in machine learning by framing the learning problem within the context of lossy compression and applying finite blocklength analysis. In our…

Beyond $L_2$: Generalizing Abductive Latent Explanations to Diverse Prototype-Based Architectures

2026-08-17 · Jules Soria, Alban Grastien, Romain Xu-Darme, Julien Girard-Satabin 외 arxiv

Prototype-based neural networks are hailed as interpretable-by-design architectures. Recently, Abductive Latent Explanations (ALE) were introduced to provide formal, mathematically guaranteed explanations that leverage t…

A Theory of Consciousness from a Theoretical Computer Science Perspective: Insights from the Conscious Turing Machine

2021-07-29 · Lenore Blum, Manuel Blum

The quest to understand consciousness, once the purview of philosophers and theologians, is now actively pursued by scientists of many stripes. We examine consciousness from the perspective of theoretical computer scienc…