paper-with-me

홈 › Papers

Deeply-Debiased Off-Policy Interval Estimation

2021-05-10 · Chengchun Shi, Runzhe Wan, Victor Chernozhukov, Rui Song

Off-policy evaluation learns a target policy's value with a historical dataset generated by a different behavior policy. In addition to a point estimate, many applications would benefit significantly from having a confidence interval (CI) that quantifies the uncertainty of the point estimate. In this paper, we propose a novel deeply-debiasing procedure to construct an efficient, robust, and flexible CI on a target policy's value. Our method is justified by theoretical results and numerical experiments. A Python implementation of the proposed procedure is available at https://github.com/RunzheStat/D2OPE.

📄 PDF Abstract BibTeX arXiv:2105.04646

Code (1)

RunzheStat/D2OPE 공식 구현 tf

Tasks

Off-policy evaluation

Similar Papers 제목 키워드 기반

ScoreMatchingRiesz: Score Matching for Debiased Machine Learning and Policy Path Estimation

2025-12-23 · Masahiro Kato arxiv

We propose ScoreMatchingRiesz, a family of Riesz representer estimators based on score matching. The Riesz representer is a key nuisance component in debiased machine learning, enabling $\sqrt{n}$-consistent and asymptot…

Automatic doubly robust inference for linear functionals via calibrated debiased machine learning

2024-11-05 · Lars van der Laan, Alex Luedtke, Marco Carone

In causal inference, many estimands of interest can be expressed as a linear functional of the outcome regression function; this includes, for example, average causal effects of static, dynamic and stochastic interventio…

Causal Inferenceregression

Double/Debiased Machine Learning for Dynamic Treatment Effects via g-Estimation

2020-02-17 · Greg Lewis, Vasilis Syrgkanis

We consider the estimation of treatment effects in settings when multiple treatments are assigned over time and treatments can have a causal effect on future outcomes or the state of the treated unit. We propose an exten…

BIG-bench Machine LearningModel SelectionOff-policy evaluation

Triple/Debiased Lasso for Statistical Inference of Conditional Average Treatment Effects

2024-03-05 · Masahiro Kato

This study investigates the estimation and the statistical inference about Conditional Average Treatment Effects (CATEs), which have garnered attention as a metric representing individualized causal effects. In our data-…

regression

Uncertainty Quantification for Demand Prediction in Contextual Dynamic Pricing

2020-03-16 · Yining Wang, Xi Chen, Xiangyu Chang, Dongdong Ge

Data-driven sequential decision has found a wide range of applications in modern operations management, such as dynamic pricing, inventory control, and assortment optimization. Most existing research on data-driven seque…

Assortment OptimizationManagementUncertainty Quantificationvalid