paper-with-me

홈 › Papers

Evaluating Counterfactual Explanation Methods on Incomplete Inputs

2026-04-09 · Francesco Leofante, Daniel Neider, Mustafa Yalçıner arxiv

Existing algorithms for generating Counterfactual Explanations (CXs) for Machine Learning (ML) typically assume fully specified inputs. However, real-world data often contains missing values, and the impact of these incomplete inputs on the performance of existing CX methods remains unexplored. To address this gap, we systematically evaluate recent CX generation methods on their ability to provide valid and plausible counterfactuals when inputs are incomplete. As part of this investigation, we hypothesize that robust CX generation methods will be better suited to address the challenge of providing valid and plausible counterfactuals when inputs are incomplete. Our findings reveal that while robust CX methods achieve higher validity than non-robust ones, all methods struggle to find valid counterfactuals. These results motivate the need for new CX methods capable of handling incomplete inputs.

📄 PDF Abstract BibTeX arXiv:2604.08004

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Faithfulness Tests for Natural Language Explanations

2023-05-29 · Pepa Atanasova, Oana-Maria Camburu, Christina Lioma, Thomas Lukasiewicz 외

Explanations of neural models aim to reveal a model's decision-making process for its predictions. However, recent work shows that current methods giving explanations such as saliency maps or counterfactuals can be misle…

counterfactualDecision Making

Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments

2026-08-17 · Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks arxiv

Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In this work, we evaluate explanations thr…

Counterfactuals As a Means for Evaluating Faithfulness of Attribution Methods in Autoregressive Language Models

2024-08-21 · Sepehr Kamahi, Yadollah Yaghoobzadeh

Despite the widespread adoption of autoregressive language models, explainability evaluation research has predominantly focused on span infilling and masked language models. Evaluating the faithfulness of an explanation …

counterfactualDecision MakingFeature ImportanceLanguage Modelling

Counterfactual Explanation with Multi-Agent Reinforcement Learning for Drug Target Prediction

2021-03-24 · Tri Minh Nguyen, Thomas P Quinn, Thin Nguyen, Truyen Tran

Motivation: Many high-performance DTA models have been proposed, but they are mostly black-box and thus lack human interpretability. Explainable AI (XAI) can make DTA models more trustworthy, and can also enable scientis…

counterfactualCounterfactual ExplanationExplainable Artificial Intelligence (XAI)Multi-agent Reinforcement Learning+2

A Comparative Analysis of Counterfactual Explanation Methods for Text Classifiers

2024-11-04 · Stephen McAleese, Mark Keane

Counterfactual explanations can be used to interpret and debug text classifiers by producing minimally altered text inputs that change a classifier's output. In this work, we evaluate five methods for generating counterf…

counterfactualCounterfactual Explanationvalid