paper-with-me

Papers

High-Dimensional Feature Selection for Sample Efficient Treatment Effect Estimation

2020-11-03 · Kristjan Greenewald, Dmitriy Katz-Rogozhnikov, Karthik Shanmugam

The estimation of causal treatment effects from observational data is a fundamental problem in causal inference. To avoid bias, the effect estimator must control for all confounders. Hence practitioners often collect data for as many covariates as possible to raise the chances of including the relevant confounders. While this addresses the bias, this has the side effect of significantly increasing the number of data samples required to accurately estimate the effect due to the increased dimensionality. In this work, we consider the setting where out of a large number of covariates $X$ that satisfy strong ignorability, an unknown sparse subset $S$ is sufficient to include to achieve zero bias, i.e. $c$-equivalent to $X$. We propose a common objective function involving outcomes across treatment cohorts with nonconvex joint sparsity regularization that is guaranteed to recover $S$ with high probability under a linear outcome model for $Y$ and subgaussian covariates for each of the treatment cohort. This improves the effect estimation sample complexity so that it scales with the cardinality of the sparse subset $S$ and $\log |X|$, as opposed to the cardinality of the full set $X$. We validate our approach with experiments on treatment effect estimation.

📄 PDF Abstract BibTeX arXiv:2011.01979

Code (0)

등록된 구현이 없습니다.

Tasks

Causal Inferencefeature selectionVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

Differentiable Pareto-Smoothed Weighting for High-Dimensional Heterogeneous Treatment Effect Estimation

2024-04-26 · Yoichi Chikahara, Kansei Ushiyama

There is a growing interest in estimating heterogeneous treatment effects across individuals using their high-dimensional feature attributes. Achieving high performance in such high-dimensional heterogeneous treatment ef…

Heterogeneous Treatment Effect EstimationRepresentation LearningSelection bias

Lee Bounds with a Continuous Treatment in Sample Selection

2024-11-06 · Ying-Ying Lee, Chu-An Liu

We study causal inference in sample selection models where a continuous or multivalued treatment affects both outcomes and their observability (e.g., employment or survey responses). We generalized the widely used Lee (2…

Causal InferenceSelection bias

Double machine learning for sample selection models

2020-11-30 · Michela Bia, Martin Huber, Lukáš Lafférs

This paper considers the evaluation of discretely distributed treatments when outcomes are only observed for a subpopulation due to sample selection or outcome attrition. For identification, we combine a selection-on-obs…

BIG-bench Machine Learning

Generalized Kernel Ridge Regression for Causal Inference with Missing-at-Random Sample Selection

2021-11-09 · Rahul Singh

I propose kernel ridge regression estimators for nonparametric dose response curves and semiparametric treatment effects in the setting where an analyst has access to a selected sample rather than a random sample; only f…

Causal Inferencecounterfactualregression

Feature Dimensionality Outweighs Model Complexity in Breast Cancer Subtype Classification Using TCGA-BRCA Gene Expression Data

2026-05-07 · Meena Al Hasani arxiv

Accurate classification of breast cancer subtypes from gene expression data is critical for diagnosis and treatment selection. However, such datasets are characterized by high dimensionality and limited sample size, posi…