paper-with-me

홈 › Papers

Bias and high-dimensional adjustment in observational studies of peer effects

2017-06-14 · Dean Eckles, Eytan Bakshy

Peer effects, in which the behavior of an individual is affected by the behavior of their peers, are posited by multiple theories in the social sciences. Other processes can also produce behaviors that are correlated in networks and groups, thereby generating debate about the credibility of observational (i.e. nonexperimental) studies of peer effects. Randomized field experiments that identify peer effects, however, are often expensive or infeasible. Thus, many studies of peer effects use observational data, and prior evaluations of causal inference methods for adjusting observational data to estimate peer effects have lacked an experimental "gold standard" for comparison. Here we show, in the context of information and media diffusion on Facebook, that high-dimensional adjustment of a nonexperimental control group (677 million observations) using propensity score models produces estimates of peer effects statistically indistinguishable from those from using a large randomized experiment (220 million observations). Naive observational estimators overstate peer effects by 320% and commonly used variables (e.g., demographics) offer little bias reduction, but adjusting for a measure of prior behaviors closely related to the focal behavior reduces bias by 91%. High-dimensional models adjusting for over 3,700 past behaviors provide additional bias reduction, such that the full model reduces bias by over 97%. This experimental evaluation demonstrates that detailed records of individuals' past behavior can improve studies of social influence, information diffusion, and imitation; these results are encouraging for the credibility of some studies but also cautionary for studies of rare or new behaviors. More generally, these results show how large, high-dimensional data sets and statistical learning techniques can be used to improve causal inference in the behavioral sciences.

📄 PDF Abstract BibTeX arXiv:1706.04692

Code (1)

fghjorth/vkme18

Tasks

Causal InferenceVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Causal inference Causal inference is the process of drawing a conclusion about a causal connection based on the conditions of the occurrence of an effect. The main difference between causal…

Similar Papers 제목 키워드 기반

Learning Adjustment Sets from Observational and Limited Experimental Data

2020-05-18 · Sofia Triantafillou, Gregory Cooper

Estimating causal effects from observational data is not always possible due to confounding. Identifying a set of appropriate covariates (adjustment set) and adjusting for their influence can remove confounding bias; how…

Recover Experimental Data with Selection Bias using Counterfactual Logic

2025-05-31 · Jingyang He, Shuai Wang, Ang Li

Selection bias, arising from the systematic inclusion or exclusion of certain samples, poses a significant challenge to the validity of causal inference. While Bareinboim et al. introduced methods for recovering unbiased…

Causal InferencecounterfactualSelection bias

Efficient adjustment for complex covariates: Gaining efficiency with DOPE

2024-02-20 · Alexander Mangulad Christgau, Niels Richard Hansen

Covariate adjustment is a ubiquitous method used to estimate the average treatment effect (ATE) from observational data. Assuming a known graphical structure of the data generating model, recent results give graphical cr…

High Dimensional Causal Inference with Variational Backdoor Adjustment

2023-10-09 · Daniel Israel, Aditya Grover, Guy Van Den Broeck

Backdoor adjustment is a technique in causal inference for estimating interventional quantities from purely observational data. For example, in medical settings, backdoor adjustment can be used to control for confounding…

Causal InferenceVariational Inference

Debiased maximum-likelihood estimators for hazard ratios under kernel-based machine-learning adjustment

2025-07-23 · Takashi Hayakawa, Satoshi Asai arxiv

Previous studies have shown that hazard ratios between treatment groups estimated with the Cox model are uninterpretable because the unspecified baseline hazard of the model fails to identify temporal change in the risk …

Causal Inference