paper-with-me

Papers

Generally-Occurring Model Change for Robust Counterfactual Explanations

2024-07-16 · Ao Xu, Tieru Wu

With the increasing impact of algorithmic decision-making on human lives, the interpretability of models has become a critical issue in machine learning. Counterfactual explanation is an important method in the field of interpretable machine learning, which can not only help users understand why machine learning models make specific decisions, but also help users understand how to change these decisions. Naturally, it is an important task to study the robustness of counterfactual explanation generation algorithms to model changes. Previous literature has proposed the concept of Naturally-Occurring Model Change, which has given us a deeper understanding of robustness to model change. In this paper, we first further generalize the concept of Naturally-Occurring Model Change, proposing a more general concept of model parameter changes, Generally-Occurring Model Change, which has a wider range of applicability. We also prove the corresponding probabilistic guarantees. In addition, we consider a more specific problem, data set perturbation, and give relevant theoretical results by combining optimization theory.

📄 PDF Abstract BibTeX arXiv:2407.11426

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualCounterfactual ExplanationDecision MakingExplanation GenerationInterpretable Machine Learningmodel

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with Counterfactual Experiments

2026-08-17 · Adam Karvonen, Euan Ong, Subhash Kantamneni, Samuel Marks arxiv

Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In this work, we evaluate explanations thr…

Robust Counterfactual Explanations for Neural Networks With Probabilistic Guarantees

2023-05-19 · Faisal Hamman, Erfaun Noorani, Saumitra Mishra, Daniele Magazzeni 외

There is an emerging interest in generating robust counterfactual explanations that would remain valid if the model is updated or changed even slightly. Towards finding robust counterfactuals, existing literature often a…

counterfactualvalid

Explainable Anomaly Detection: Counterfactual driven What-If Analysis

2024-08-21 · Logan Cummins, Alexander Sommers, Sudip Mittal, Shahram Rahimi 외

There exists three main areas of study inside of the field of predictive maintenance: anomaly detection, fault diagnosis, and remaining useful life prediction. Notably, anomaly detection alerts the stakeholder that an an…

Anomaly DetectioncounterfactualExplainable artificial intelligenceFault Diagnosis

Evaluating Robustness of Counterfactual Explanations

2021-03-03 · André Artelt, Valerie Vaquet, Riza Velioglu, Fabian Hinder 외

Transparency is a fundamental requirement for decision making systems when these should be deployed in the real world. It is usually achieved by providing explanations of the system's behavior. A prominent and intuitive …

counterfactualDecision MakingFairness

If Only We Had Better Counterfactual Explanations: Five Key Deficits to Rectify in the Evaluation of Counterfactual XAI Techniques

2021-02-26 · Mark T Keane, Eoin M Kenny, Eoin Delaney, Barry Smyth

In recent years, there has been an explosion of AI research on counterfactual explanations as a solution to the problem of eXplainable AI (XAI). These explanations seem to offer technical, psychological and legal benefit…

counterfactualCounterfactual ExplanationExplainable Artificial Intelligence (XAI)Survey