paper-with-me

홈 › Papers

Improving Bias Correction Standards by Quantifying its Effects on Treatment Outcomes

2024-07-20 · Alexandre Abraham, Andrés Hoyos Idrobo

With the growing access to administrative health databases, retrospective studies have become crucial evidence for medical treatments. Yet, non-randomized studies frequently face selection biases, requiring mitigation strategies. Propensity score matching (PSM) addresses these biases by selecting comparable populations, allowing for analysis without further methodological constraints. However, PSM has several drawbacks. Different matching methods can produce significantly different Average Treatment Effects (ATE) for the same task, even when meeting all validation criteria. To prevent cherry-picking the best method, public authorities must involve field experts and engage in extensive discussions with researchers. To address this issue, we introduce a novel metric, A2A, to reduce the number of valid matches. A2A constructs artificial matching tasks that mirror the original ones but with known outcomes, assessing each matching method's performance comprehensively from propensity estimation to ATE estimation. When combined with Standardized Mean Difference, A2A enhances the precision of model selection, resulting in a reduction of up to 50% in ATE estimation errors across synthetic tasks and up to 90% in predicted ATE variability across both synthetic and real-world datasets. To our knowledge, A2A is the first metric capable of evaluating outcome correction accuracy using covariates not involved in selection. Computing A2A requires solving hundreds of PSMs, we therefore automate all manual steps of the PSM pipeline. We integrate PSM methods from Python and R, our automated pipeline, a new metric, and reproducible experiments into popmatch, our new Python package, to enhance reproducibility and accessibility to bias correction methods.

📄 PDF Abstract BibTeX arXiv:2407.14861

Code (0)

등록된 구현이 없습니다.

Tasks

Face SelectionModel Selection

Similar Papers 제목 키워드 기반

Automatic Debiased Machine Learning for Dynamic Treatment Effects and General Nested Functionals

2022-03-25 · Victor Chernozhukov, Whitney Newey, Rahul Singh, Vasilis Syrgkanis

We extend the idea of automated debiased machine learning to the dynamic treatment regime and more generally to nested functionals. We show that the multiply robust formula for the dynamic treatment regime with discrete …

BIG-bench Machine LearningDiscrete Choice Models

Inference on effect size after multiple hypothesis testing

2025-03-28 · Andreas Dzemski, Ryo Okui, Wenjie Wang

Significant treatment effects are often emphasized when interpreting and summarizing empirical findings in studies that estimate multiple, possibly many, treatment effects. Under this kind of selective reporting, convent…

valid

Detecting and Mitigating Group Bias in Heterogeneous Treatment Effects

2026-02-23 · Joel Persson, Jurriën Bakker, Dennis Bohle, Stefan Feuerriegel 외 arxiv

Heterogeneous treatment effects (HTEs) are increasingly estimated using machine learning models that produce highly personalized predictions of treatment effects. In practice, however, predicted treatment effects are rar…

Dynamic Biases of Static Panel Data Estimators

2024-10-21 · Sylvia Klosin

This paper identifies an important bias - termed dynamic bias - in fixed effects panel estimators that arises when dynamic feedback is ignored in the estimating equation. Dynamic feedback occurs if past outcomes impact c…

Optimal selection of the number of control units in kNN algorithm to estimate average treatment effects

2020-08-14

We propose a simple approach to optimally select the number of control units in k nearest neighbors (kNN) algorithm focusing in minimizing the mean squared error for the average treatment effects. Our approach is non-par…