paper-with-me

Papers

Double machine learning for sample selection models

2020-11-30 · Michela Bia, Martin Huber, Lukáš Lafférs

This paper considers the evaluation of discretely distributed treatments when outcomes are only observed for a subpopulation due to sample selection or outcome attrition. For identification, we combine a selection-on-observables assumption for treatment assignment with either selection-on-observables or instrumental variable assumptions concerning the outcome attrition/sample selection process. We also consider dynamic confounding, meaning that covariates that jointly affect sample selection and the outcome may (at least partly) be influenced by the treatment. To control in a data-driven way for a potentially high dimensional set of pre- and/or post-treatment covariates, we adapt the double machine learning framework for treatment evaluation to sample selection problems. We make use of (a) Neyman-orthogonal, doubly robust, and efficient score functions, which imply the robustness of treatment effect estimation to moderate regularization biases in the machine learning-based estimation of the outcome, treatment, or sample selection models and (b) sample splitting (or cross-fitting) to prevent overfitting bias. We demonstrate that the proposed estimators are asymptotically normal and root-n consistent under specific regularity conditions concerning the machine learners and investigate their finite sample properties in a simulation study. We also apply our proposed methodology to the Job Corps data for evaluating the effect of training on hourly wages which are only observed conditional on employment. The estimator is available in the causalweight package for the statistical software R.

📄 PDF Abstract BibTeX arXiv:2012.00745

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Automatic debiased machine learning and sensitivity analysis for sample selection models

2026-01-13 · Jakob Bjelac, Victor Chernozhukov, Phil-Adrian Klotz, Jannis Kueck 외 arxiv

In this paper, we extend the Riesz representation framework to causal inference under sample selection, where both treatment assignment and outcome observability are non-random. Formulating the problem in terms of a Ries…

Causal Inference

DoubleEnsemble: A New Ensemble Method Based on Sample Reweighting and Feature Selection for Financial Data Analysis

2020-10-03 · Chuheng Zhang, Yuanqi Li, Xi Chen, Yifei Jin 외

Modern machine learning models (such as deep neural networks and boosting decision tree models) have become increasingly popular in financial market prediction, due to their superior capacity to extract complex non-linea…

BIG-bench Machine Learningfeature selection

Lee Bounds with a Continuous Treatment in Sample Selection

2024-11-06 · Ying-Ying Lee, Chu-An Liu

We study causal inference in sample selection models where a continuous or multivalued treatment affects both outcomes and their observability (e.g., employment or survey responses). We generalized the widely used Lee (2…

Causal InferenceSelection bias

Understanding the double descent curve in Machine Learning

2022-11-18 · Luis Sa-Couto, Jose Miguel Ramos, Miguel Almeida, Andreas Wichert

The theory of bias-variance used to serve as a guide for model selection when applying Machine Learning algorithms. However, modern practice has shown success with over-parameterized models that were expected to overfit …

Model Selection

Evaluating (weighted) dynamic treatment effects by double machine learning

2020-12-01 · Hugo Bodory, Martin Huber, Lukáš Lafférs

We consider evaluating the causal effects of dynamic treatments, i.e. of multiple treatment sequences in various periods, based on double machine learning to control for observed, time-varying covariates in a data-driven…

BIG-bench Machine Learning