paper-with-me

홈 › Papers

Rigorous Probabilistic Guarantees for Robust Counterfactual Explanations

2024-07-10 · Luca Marzari, Francesco Leofante, Ferdinando Cicalese, Alessandro Farinelli

We study the problem of assessing the robustness of counterfactual explanations for deep learning models. We focus on $\textit{plausible model shifts}$ altering model parameters and propose a novel framework to reason about the robustness property in this setting. To motivate our solution, we begin by showing for the first time that computing the robustness of counterfactuals with respect to plausible model shifts is NP-complete. As this (practically) rules out the existence of scalable algorithms for exactly computing robustness, we propose a novel probabilistic approach which is able to provide tight estimates of robustness with strong guarantees while preserving scalability. Remarkably, and differently from existing solutions targeting plausible model shifts, our approach does not impose requirements on the network to be analyzed, thus enabling robustness analysis on a wider range of architectures. Experiments on four binary classification datasets indicate that our method improves the state of the art in generating robust explanations, outperforming existing methods on a range of metrics.

📄 PDF Abstract BibTeX arXiv:2407.07482

Code (1)

lmarza/apas 공식 구현 tf

Tasks

Binary Classificationcounterfactual

Methods 이 논문이 사용한 방법론

Focus 설명 없음
Counterfactuals 설명 없음

Similar Papers 제목 키워드 기반

CONFEX: Uncertainty-Aware Counterfactual Explanations with Conformal Guarantees

2025-10-22 · Aman Bilkhoo, Mehran Hosseini, Milad Kazemi, Nicola Paoletti arxiv

Counterfactual explanations (CFXs) provide human-understandable justifications for model predictions, enabling actionable recourse and enhancing interpretability. To be reliable, CFXs must avoid regions of high predictiv…

Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change

2024-08-09 · Ignacy Stępka, Mateusz Lango, Jerzy Stefanowski

Counterfactual explanations (CFEs) guide users on how to adjust inputs to machine learning models to achieve desired outputs. While existing research primarily addresses static scenarios, real-world applications often in…

counterfactual

Provably Robust Bayesian Counterfactual Explanations under Model Changes

2026-01-23 · Jamie Duell, Xiuyi Fan arxiv

Counterfactual explanations (CEs) offer interpretable insights into machine learning predictions by answering ``what if?" questions. However, in real-world settings where models are frequently updated, existing counterfa…

Don't Explain Noise: Robust Counterfactuals for Randomized Ensembles

2022-05-27 · Alexandre Forel, Axel Parmentier, Thibaut Vidal

Counterfactual explanations describe how to modify a feature vector in order to flip the outcome of a trained classifier. Obtaining robust counterfactual explanations is essential to provide valid algorithmic recourse an…

counterfactualvalid

Local and Regional Counterfactual Rules: Summarized and Robust Recourses

2022-09-29 · Salim I. Amoukou, Nicolas J. B Brunel

Counterfactual Explanations (CE) face several unresolved challenges, such as ensuring stability, synthesizing multiple CEs, and providing plausibility and sparsity guarantees. From a more practical point of view, recent …

counterfactual