Machine-Learning Tests for Effects on Multiple Outcomes
In this paper we present tools for applied researchers that re-purpose off-the-shelf methods from the computer-science field of machine learning to create a "discovery engine" for data from randomized controlled trials (RCTs). The applied problem we seek to solve is that economists invest vast resources into carrying out RCTs, including the collection of a rich set of candidate outcome measures. But given concerns about inference in the presence of multiple testing, economists usually wind up exploring just a small subset of the hypotheses that the available data could be used to test. This prevents us from extracting as much information as possible from each RCT, which in turn impairs our ability to develop new theories or strengthen the design of policy interventions. Our proposed solution combines the basic intuition of reverse regression, where the dependent variable of interest now becomes treatment assignment itself, with methods from machine learning that use the data themselves to flexibly identify whether there is any function of the outcomes that predicts (or has signal about) treatment group status. This leads to correctly-sized tests with appropriate $p$-values, which also have the important virtue of being easy to implement in practice. One open challenge that remains with our work is how to meaningfully interpret the signal that these methods find.
Code (0)
등록된 구현이 없습니다.
Tasks
BIG-bench Machine LearningSimilar Papers 제목 키워드 기반
Evaluating Policies Early in a Pandemic: Bounding Policy Effects with Nonrandomly Missing Data
During the early part of the Covid-19 pandemic, national and local governments introduced a number of policies to combat the spread of Covid-19. In this paper, we propose a new approach to bound the effects of such early…
Estimating Nonlinear Network Data Models with Fixed Effects
I introduce a new method for bias correction of dyadic models with agent-specific fixed effects, including the dyadic link formation model with homophily and degree heterogeneity. The proposed approach uses a jackknife p…
counterfactualTesting for homogeneous treatment effects in linear and nonparametric instrumental variable models
The hypothesis of homogeneous treatment effects is central to the instrumental variables literature. This assumption signifies that treatment effects are constant across all subjects. It allows to interpret instrumental …
Identifying Heterogeneous Treatment Effects in Multiple Outcomes using Joint Confidence Intervals
Heterogeneous treatment effects (HTEs) are commonly identified during randomized controlled trials (RCTs). Identifying subgroups of patients with similar treatment effects is of high interest in clinical research to adva…
Doubly Robust Kernel Statistics for Testing Distributional Treatment Effects
With the widespread application of causal inference, it is increasingly important to have tools which can test for the presence of causal effects in a diverse array of circumstances. In this vein we focus on the problem …
Causal InferencecounterfactualOff-policy evaluation