paper-with-me

홈 › Papers

Post-Hoc Explanations Fail to Achieve their Purpose in Adversarial Contexts

2022-01-25 · Sebastian Bordt, Michèle Finck, Eric Raidl, Ulrike Von Luxburg

Existing and planned legislation stipulates various obligations to provide information about machine learning algorithms and their functioning, often interpreted as obligations to "explain". Many researchers suggest using post-hoc explanation algorithms for this purpose. In this paper, we combine legal, philosophical and technical arguments to show that post-hoc explanation algorithms are unsuitable to achieve the law's objectives. Indeed, most situations where explanations are requested are adversarial, meaning that the explanation provider and receiver have opposing interests and incentives, so that the provider might manipulate the explanation for her own ends. We show that this fundamental conflict cannot be resolved because of the high degree of ambiguity of post-hoc explanations in realistic application scenarios. As a consequence, post-hoc explanation algorithms are unsuitable to achieve the transparency objectives inherent to the legal norms. Instead, there is a need to more explicitly discuss the objectives underlying "explainability" obligations as these can often be better achieved through other mechanisms. There is an urgent need for a more open and honest discussion regarding the potential and limitations of post-hoc explanations in adversarial contexts, in particular in light of the current negotiations of the European Union's draft Artificial Intelligence Act.

📄 PDF Abstract BibTeX arXiv:2201.10295

Code (1)

tml-tuebingen/facct-post-hoc 공식 구현

Similar Papers 제목 키워드 기반

ExpProof : Operationalizing Explanations for Confidential Models with ZKPs

2025-02-06 · Chhavi Yadav, Evan Monroe Laufer, Dan Boneh, Kamalika Chaudhuri

In principle, explanations are intended as a way to increase trust in machine learning models and are often obligated by regulations. However, many circumstances where these are demanded are adversarial in nature, meanin…

Robust Ante-hoc Graph Explainer using Bilevel Optimization

2023-05-25 · Kha-Dinh Luong, Mert Kosan, Arlei Lopes da Silva, Ambuj Singh

Explaining the decisions made by machine learning models for high-stakes applications is critical for increasing transparency and guiding improvements to these decisions. This is particularly true in the case of models f…

AttributeBilevel OptimizationGraph Classification

Unsynthesizable Cores - Minimal Explanations for Unsynthesizable High-Level Robot Behaviors

2014-09-04 · Vasumathi Raman, Hadas Kress-Gazit

With the increasing ubiquity of multi-capable, general-purpose robots arises the need for enabling non-expert users to command these robots to perform complex high-level tasks. To this end, high-level robot control has s…

Vocal Bursts Intensity Prediction

Provably Better Explanations with Optimized Aggregation of Feature Attributions

2024-06-07 · Thomas Decker, Ananta R. Bhattarai, Jindong Gu, Volker Tresp 외

Using feature attributions for post-hoc explanations is a common practice to understand and verify the predictions of opaque machine learning models. Despite the numerous techniques available, individual methods often pr…

Beyond Accuracy, SHAP, and Anchors -- On the difficulty of designing effective end-user explanations

2025-01-28 · Zahra Abba Omar, Nadia Nahar, Jacob Tjaden, Inès M. Gilles 외

Modern machine learning produces models that are impossible for users or developers to fully understand -- raising concerns about trust, oversight and human dignity. Transparency and explainability methods aim to provide…