Data-driven Policy Learning for Continuous Treatments
This paper studies policy learning for continuous treatments from observational data. Continuous treatments present more significant challenges than discrete ones because population welfare may need nonparametric estimation, and policy space may be infinite-dimensional and may satisfy shape restrictions. We propose to approximate the policy space with a sequence of finite-dimensional spaces and, for any given policy, obtain the empirical welfare by applying the kernel method. We consider two cases: known and unknown propensity scores. In the latter case, we allow for machine learning of the propensity score and modify the empirical welfare to account for the effect of machine learning. The learned policy maximizes the empirical welfare or the modified empirical welfare over the approximating space. In both cases, we modify the penalty algorithm proposed in \cite{mbakop2021model} to data-automate the tuning parameters (i.e., bandwidth and dimension of the approximating space) and establish an oracle inequality for the welfare regret.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Causal Modeling of Policy Interventions From Sequences of Treatments and Outcomes
A treatment policy defines when and what treatments are applied to affect some outcome of interest. Data-driven decision-making requires the ability to predict what happens if a policy is changed. Existing methods that p…
counterfactualDecision MakingGaussian ProcessesPoint Processes+2Policy Evaluation and Optimization with Continuous Treatments
We study the problem of policy evaluation and learning from batched contextual bandit data when treatments are continuous, going beyond previous work on discrete treatments. Previous work for discrete treatment/action sp…
Deep Jump Learning for Off-Policy Evaluation in Continuous Treatment Settings
We consider off-policy evaluation (OPE) in continuous treatment settings, such as personalized dose-finding. In OPE, one aims to estimate the mean outcome under a new treatment decision rule using historical data generat…
Change Point DetectionOff-policy evaluationQ-LearningDistributionally Robust Policy Evaluation and Learning for Continuous Treatment with Observational Data
Using offline observational data for policy evaluation and learning allows decision-makers to evaluate and learn a policy that connects characteristics and interventions. Most existing literature has focused on either di…
Changes-In-Changes For Discrete Treatment
This paper generalizes the changes-in-changes (CIC) model to handle discrete treatments with more than two categories, extending the binary case of Athey and Imbens (2006). While the original CIC model is well-suited for…