paper-with-me

Papers

Variable importance without impossible data

2022-05-31 · Masayoshi Mase, Art B. Owen, Benjamin B. Seiler

The most popular methods for measuring importance of the variables in a black box prediction algorithm make use of synthetic inputs that combine predictor variables from multiple subjects. These inputs can be unlikely, physically impossible, or even logically impossible. As a result, the predictions for such cases can be based on data very unlike any the black box was trained on. We think that users cannot trust an explanation of the decision of a prediction algorithm when the explanation uses such values. Instead we advocate a method called Cohort Shapley that is grounded in economic game theory and unlike most other game theoretic methods, it uses only actually observed data to quantify variable importance. Cohort Shapley works by narrowing the cohort of subjects judged to be similar to a target subject on one or more features. We illustrate it on an algorithmic fairness problem where it is essential to attribute importance to protected variables that the model was not trained on.

📄 PDF Abstract BibTeX arXiv:2205.15750

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeFairness

Similar Papers 제목 키워드 기반

Explaining black box decisions by Shapley cohort refinement

2019-11-01 · Masayoshi Mase, Art B. Owen, Benjamin Seiler

We introduce a variable importance measure to quantify the impact of individual input variables to a black box function. Our measure is based on the Shapley value from cooperative game theory. Many measures of variable i…

Cohort Shapley value for algorithmic fairness

2021-05-15 · Masayoshi Mase, Art B. Owen, Benjamin B. Seiler

Cohort Shapley value is a model-free method of variable importance grounded in game theory that does not use any unobserved and potentially impossible feature combinations. We use it to evaluate algorithmic fairness, usi…

AttributeFairness

Restricted Hidden Cardinality Constraints in Causal Models

2021-09-13 · Beata Zjawin, Elie Wolfe, Robert W. Spekkens

Causal models with unobserved variables impose nontrivial constraints on the distributions over the observed variables. When a common cause of two variables is unobserved, it is impossible to uncover the causal relation …

The Rashomon Importance Distribution: Getting RID of Unstable, Single Model-based Variable Importance

2023-09-24 · NeurIPS 2023 11 · Jon Donnelly, Srikar Katta, Cynthia Rudin, Edward P. Browne

Quantifying variable importance is essential for answering high-stakes questions in fields like genetics, public policy, and medicine. Current methods generally calculate variable importance for a given model trained on …

Weakly supervised causal representation learning

2022-03-30 · Johann Brehmer, Pim de Haan, Phillip Lippe, Taco Cohen

Learning high-level causal representations together with a causal model from unstructured low-level data such as pixels is impossible from observational data alone. We prove under mild assumptions that this representatio…

Representation Learning