paper-with-me

홈 › Papers

Controlling for Unobserved Confounding with Large Language Model Classification of Patient Smoking Status

2024-11-05 · Samuel Lee, Zach Wood-Doughty

Causal understanding is a fundamental goal of evidence-based medicine. When randomization is impossible, causal inference methods allow the estimation of treatment effects from retrospective analysis of observational data. However, such analyses rely on a number of assumptions, often including that of no unobserved confounding. In many practical settings, this assumption is violated when important variables are not explicitly measured in the clinical record. Prior work has proposed to address unobserved confounding with machine learning by imputing unobserved variables and then correcting for the classifier's mismeasurement. When such a classifier can be trained and the necessary assumptions are met, this method can recover an unbiased estimate of a causal effect. However, such work has been limited to synthetic data, simple classifiers, and binary variables. This paper extends this methodology by using a large language model trained on clinical notes to predict patients' smoking status, which would otherwise be an unobserved confounder. We then apply a measurement error correction on the categorical predicted smoking status to estimate the causal effect of transthoracic echocardiography on mortality in the MIMIC dataset.

📄 PDF Abstract BibTeX arXiv:2411.03004

Code (0)

등록된 구현이 없습니다.

Tasks

Causal InferenceLanguage ModelingLanguage ModellingLarge Language Model

Methods 이 논문이 사용한 방법론

Causal inference Causal inference is the process of drawing a conclusion about a causal connection based on the conditions of the occurrence of an effect. The main difference between causal…

Similar Papers 제목 키워드 기반

Controlling for Unobserved Confounds in Classification Using Correlational Constraints

2017-03-05 · Virgile Landeiro, Aron Culotta

As statistical classifiers become integrated into real-world applications, it is important to consider not only their accuracy but also their robustness to changes in the data distribution. In this paper, we consider the…

ClassificationGeneral Classification

Confounding Robust Deep Reinforcement Learning: A Causal Approach

2025-10-24 · Mingxuan Li, Junzhe Zhang, Elias Bareinboim arxiv

A key task in Artificial Intelligence is learning effective policies for controlling agents in unknown environments to optimize performance measures. Off-policy learning methods, like Q-learning, allow learners to make o…

Reinforcement LearningAtari Games

Automatic Reward Shaping from Confounded Offline Data

2025-05-16 · Mingxuan Li, Junzhe Zhang, Elias Bareinboim

A key task in Artificial Intelligence is learning effective policies for controlling agents in unknown environments to optimize performance measures. Off-policy learning methods, like Q-learning, allow learners to make o…

Atari GamesDeep Reinforcement LearningQ-Learning

Causal Fairness under Unobserved Confounding: A Neural Sensitivity Framework

2023-11-30 · Maresa Schröder, Dennis Frauen, Stefan Feuerriegel

Fairness for machine learning predictions is widely required in practice for legal, ethical, and societal reasons. Existing work typically focuses on settings without unobserved confounding, even though unobserved confou…

FairnessSensitivity

Hidden yet quantifiable: A lower bound for confounding strength using randomized trials

2023-12-06 · Piersilvio De Bartolomeis, Javier Abad, Konstantin Donhauser, Fanny Yang

In the era of fast-paced precision medicine, observational studies play a major role in properly evaluating new treatments in clinical practice. Yet, unobserved confounding can significantly compromise causal conclusions…

valid