paper-with-me

홈 › Papers

Logic Explanation of AI Classifiers by Categorical Explaining Functors

2025-03-20 · Stefano Fioravanti, Francesco Giannini, Paolo Frazzetto, Fabio Zanasi, Pietro Barbiero

The most common methods in explainable artificial intelligence are post-hoc techniques which identify the most relevant features used by pretrained opaque models. Some of the most advanced post hoc methods can generate explanations that account for the mutual interactions of input features in the form of logic rules. However, these methods frequently fail to guarantee the consistency of the extracted explanations with the model's underlying reasoning. To bridge this gap, we propose a theoretically grounded approach to ensure coherence and fidelity of the extracted explanations, moving beyond the limitations of current heuristic-based approaches. To this end, drawing from category theory, we introduce an explaining functor which structurally preserves logical entailment between the explanation and the opaque model's reasoning. As a proof of concept, we validate the proposed theoretical constructions on a synthetic benchmark verifying how the proposed approach significantly mitigates the generation of contradictory or unfaithful explanations.

📄 PDF Abstract BibTeX arXiv:2503.16203

Code (0)

등록된 구현이 없습니다.

Tasks

Explainable artificial intelligence

Methods 이 논문이 사용한 방법론

HOC 설명 없음

Similar Papers 제목 키워드 기반

The Homunculus Brain and Categorical Logic

2019-02-28 · Michael Heller

The interaction between syntax (formal language) and its semantics (meanings of language) is one which has been well studied in categorical logic. The results of this particular study are employed to understand how the b…

Homomorphic Encryption of Intuitionistic Logic Proofs and Functional Programs: A Categorical Approach Inspired by Composite-Order Bilinear Groups

2025-02-26 · Ben Goertzel

We present a conceptual framework for extending homomorphic encryption beyond arithmetic or Boolean operations into the domain of intuitionistic logic proofs and, by the Curry-Howard correspondence, into the domain of ty…

The Shape of Explanations: A Topological Account of Rule-Based Explanations in Machine Learning

2023-01-22 · Brett Mullins

Rule-based explanations provide simple reasons explaining the behavior of machine learning classifiers at given points in the feature space. Several recent methods (Anchors, LORE, etc.) purport to generate rule-based exp…

Categorical Hopfield Networks

2022-01-08 · Matilde Marcolli

This paper discusses a simple and explicit toy-model example of the categorical Hopfield equations introduced in previous work of Manin and the author. These describe dynamical assignments of resources to networks, where…

Categorical Description Of Plant Morphogenesis

2017-02-11

This article presents formalistic tool for description of structural and biochemical relations between cells in the course of development of the body of plants. This is flexible formalistic space, based on the Category t…