A Double Machine Learning Approach to Combining Experimental and Observational Data
Experimental and observational studies often lack validity due to untestable assumptions. We propose a double machine learning approach to combine experimental and observational studies, allowing practitioners to test for assumption violations and estimate treatment effects consistently. Our framework tests for violations of external validity and ignorability under milder assumptions. When only one of these assumptions is violated, we provide semiparametrically efficient treatment effect estimators. However, our no-free-lunch theorem highlights the necessity of accurately identifying the violated assumption for consistent treatment effect estimation. Through comparative analyses, we show our framework's superiority over existing data fusion methods. The practical utility of our approach is further exemplified by three real-world case studies, underscoring its potential for widespread application in empirical research.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Double Robust Representation Learning for Counterfactual Prediction
Causal inference, or counterfactual prediction, is central to decision making in healthcare, policy and social sciences. To de-bias causal estimators with high-dimensional data in observational studies, recent advances s…
Causal InferencecounterfactualDecision MakingPrediction+1Double Machine Learning Methods for Estimating Average Treatment Effects: A Comparative Study
Observational cohort studies are increasingly being used for comparative effectiveness research to assess the safety of therapeutics. Recently, various doubly robust methods have been proposed for average treatment effec…
regressionCausal Markov Boundaries
Feature selection is an important problem in machine learning, which aims to select variables that lead to an optimal predictive model. In this paper, we focus on feature selection for post-intervention outcome predictio…
feature selectionAnytime-Valid Inference for Double/Debiased Machine Learning of Causal Parameters
Double (debiased) machine learning (DML) has seen widespread use in recent years for learning causal/structural parameters, in part due to its flexibility and adaptability to high-dimensional nuisance functions as well a…
validEfficient estimation of weighted cumulative treatment effects by double/debiased machine learning
In empirical studies with time-to-event outcomes, investigators often leverage observational data to conduct causal inference on the effect of exposure when randomized controlled trial data is unavailable. Model misspeci…
Causal Inference