paper-with-me

홈 › Papers

Performance of Cross-Validated Targeted Maximum Likelihood Estimation

2024-09-17 · Matthew J. Smith, Rachael V. Phillips, Camille Maringe, Miguel Angel Luque-Fernandez

Background: Advanced methods for causal inference, such as targeted maximum likelihood estimation (TMLE), require certain conditions for statistical inference. However, in situations where there is not differentiability due to data sparsity or near-positivity violations, the Donsker class condition is violated. In such situations, TMLE variance can suffer from inflation of the type I error and poor coverage, leading to conservative confidence intervals. Cross-validation of the TMLE algorithm (CVTMLE) has been suggested to improve on performance compared to TMLE in settings of positivity or Donsker class violations. We aim to investigate the performance of CVTMLE compared to TMLE in various settings. Methods: We utilised the data-generating mechanism as described in Leger et al. (2022) to run a Monte Carlo experiment under different Donsker class violations. Then, we evaluated the respective statistical performances of TMLE and CVTMLE with different super learner libraries, with and without regression tree methods. Results: We found that CVTMLE vastly improves confidence interval coverage without adversely affecting bias, particularly in settings with small sample sizes and near-positivity violations. Furthermore, incorporating regression trees using standard TMLE with ensemble super learner-based initial estimates increases bias and variance leading to invalid statistical inference. Conclusions: It has been shown that when using CVTMLE the Donsker class condition is no longer necessary to obtain valid statistical inference when using regression trees and under either data sparsity or near-positivity violations. We show through simulations that CVTMLE is much less sensitive to the choice of the super learner library and thereby provides better estimation and inference in cases where the super learner library uses more flexible candidates and is prone to overfitting.

📄 PDF Abstract BibTeX arXiv:2409.11265

Code (1)

mattyjsmith/CVTMLE 공식 구현

Tasks

Causal Inferenceregression

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

Bayesian implementation of Targeted Maximum Likelihood Estimation for uncertainty quantification in causal effect estimation

2025-07-21 · Saideep Nannapaneni, Joseph Sakaya, Kyle Caron, Pedro HM Albuquerque 외 arxiv

Robust decision making involves making decisions in the presence of uncertainty and is often used in critical domains such as healthcare, supply chains, and finance. Causality plays a crucial role in decision-making as i…

Decision Making

Efficient Targeted Maximum Likelihood Estimators for Two-Phase Design Problems

2026-02-27 · Sky Qiu, Susan Gruber, Pamela A. Shaw, Brian D. Williamson 외 arxiv

In a typical two-phase design, a random sample is drawn from the target population in phase 1, during which only a subset of variables is collected. In phase 2, a subsample of the phase-1 cohort is selected, and addition…

Fitting a deeply-nested hierarchical model to a large book review dataset using a moment-based estimator

2018-06-01 · Ningshan Zhang, Kyle Schmaus, Patrick O. Perry

We consider a particular instance of a common problem in recommender systems: using a database of book reviews to inform user-targeted recommendations. In our dataset, books are categorized into genres and sub-genres. To…

Recommendation Systems

The Impact of Regularization on High-dimensional Logistic Regression

2019-06-10 · NeurIPS 2019 12 · Fariborz Salehi, Ehsan Abbasi, Babak Hassibi

Logistic regression is commonly used for modeling dichotomous outcomes. In the classical setting, where the number of observations is much larger than the number of parameters, properties of the maximum likelihood estima…

regressionVocal Bursts Intensity Prediction

More Efficient Off-Policy Evaluation through Regularized Targeted Learning

2019-12-13 · Aurélien F. Bibaut, Ivana Malenica, Nikos Vlassis, Mark J. Van Der Laan

We study the problem of off-policy evaluation (OPE) in Reinforcement Learning (RL), where the aim is to estimate the performance of a new policy given historical data that may have been generated by a different policy, o…

Causal InferenceOff-policy evaluationReinforcement LearningReinforcement Learning (RL)