paper-with-me

홈 › Papers

A Generalized Variable Importance Metric and Estimator for Black Box Machine Learning Models

2022-12-20 · Mohammad Kaviul Anam Khan, Olli Saarela, Rafal Kustra

In this paper we define a population parameter, ``Generalized Variable Importance Metric (GVIM)'', to measure importance of predictors for black box machine learning methods, where the importance is not represented by model-based parameter. GVIM is defined for each input variable, using the true conditional expectation function, and it measures the variable's importance in affecting a continuous or a binary response. We extend previously published results to show that the defined GVIM can be represented as a function of the Conditional Average Treatment Effect (CATE) for any kind of a predictor, which gives it a causal interpretation and further justification as an alternative to classical measures of significance that are only available in simple parametric models. Extensive set of simulations using realistically complex relationships between covariates and outcomes and number of regression techniques of varying degree of complexity show the performance of our proposed estimator of the GVIM.

📄 PDF Abstract BibTeX arXiv:2212.09931

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Marginal and Conditional Importance Measures from Machine Learning Models and Their Relationship with Conditional Average Treatment Effect

2025-01-28 · Mohammad Kaviul Anam Khan, Olli Saarela, Rafal Kustra

Interpreting black-box machine learning models is challenging due to their strong dependence on data and inherently non-parametric nature. This paper reintroduces the concept of importance through "Marginal Variable Impo…

Coupled Gradient Estimators for Discrete Latent Variables

2021-06-15 · NeurIPS 2021 12 · Zhe Dong, andriy mnih, George Tucker

Training models with discrete latent variables is challenging due to the high variance of unbiased gradient estimators. While low-variance reparameterization gradients of a continuous relaxation can provide an effective …

A Semi-Parametric Bayesian Generalized Least Squares Estimator

2020-11-20 · Ruochen Wu, Melvyn Weeks

In this paper we propose a semi-parametric Bayesian Generalized Least Squares estimator. In a generic setting where each error is a vector, the parametric Generalized Least Square estimator maintains the assumption that …

High Dimensional Model Representation as a Glass Box in Supervised Machine Learning

2018-07-26 · Caleb Deen Bastian, Herschel Rabitz

Prediction and explanation are key objects in supervised machine learning, where predictive models are known as black boxes and explanatory models are known as glass boxes. Explanation provides the necessary and sufficie…

BIG-bench Machine Learning

Efficient nonparametric statistical inference on population feature importance using Shapley values

2020-06-16 · ICML 2020 1 · Brian D. Williamson, Jean Feng

The true population-level importance of a variable in a prediction task provides useful knowledge about the underlying data-generating mechanism and can help in deciding which measurements to collect in subsequent experi…

Feature ImportanceMortality Predictionvalid