paper-with-me

홈 › Papers

Reliable Off-Policy Learning for Dosage Combinations

2023-09-21 · NeurIPS 2023 11

Decision-making in personalized medicine such as cancer therapy or critical care must often make choices for dosage combinations, i.e., multiple continuous treatments. Existing work for this task has modeled the effect of multiple treatments independently, while estimating the joint effect has received little attention but comes with non-trivial challenges. In this paper, we propose a novel method for reliable off-policy learning for dosage combinations. Our method proceeds along three steps: (1) We develop a tailored neural network that estimates the individualized dose-response function while accounting for the joint effect of multiple dependent dosages. (2) We estimate the generalized propensity score using conditional normalizing flows in order to detect regions with limited overlap in the shared covariate-treatment space. (3) We present a gradient-based learning algorithm to find the optimal, individualized dosage combinations. Here, we ensure reliable estimation of the policy value by avoiding regions with limited overlap. We finally perform an extensive evaluation of our method to show its effectiveness. To the best of our knowledge, ours is the first work to provide a method for reliable off-policy learning for optimal dosage combinations.

📄 PDF Abstract BibTeX

Code (1)

jschweisthal/reliabledosagecombi 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Normalizing Flows Normalizing Flows are a method for constructing complex distributions by transforming a probability density through a series of invertible mappings. By repeatedly applying…

Similar Papers 제목 키워드 기반

Optimizing Warfarin Dosing Using Contextual Bandit: An Offline Policy Learning and Evaluation Method

2024-02-16 · Yong Huang, Charles A. Downs, Amir M. Rahmani

Warfarin, an anticoagulant medication, is formulated to prevent and address conditions associated with abnormal blood clotting, making it one of the most prescribed drugs globally. However, determining the suitable dosag…

Decision Making

Towards generalizable single-cell perturbation modeling via the Conditional Monge Gap

2025-04-11 · Alice Driessen, Benedek Harsanyi, Marianna Rapsomaniki, Jannis Born

Learning the response of single-cells to various treatments offers great potential to enable targeted therapies. In this context, neural optimal transport (OT) has emerged as a principled methodological framework because…

Predicting Dosage of Immunosuppressant Drugs After Kidney Transplantation Using Machine Learning

2023-08-22 · Kapil Panda, Anirudh Mazumder

While kidney transplants are seen as the best treatment option for patients with end-stage renal disease and kidney failure, the organ's health depends on the dosage of immunosuppressant drugs post-transplantation. Due t…

Extracting Daily Dosage from Medication Instructions in EHRs: An Automated Approach and Lessons Learned

2020-05-21 · Diwakar Mahajan, Jennifer J. Liang, Ching-Huei Tsou

Medication timelines have been shown to be effective in helping physicians visualize complex patient medication information. A key feature in many such designs is a longitudinal representation of a medication's daily dos…

Sublinear Optimal Policy Value Estimation in Contextual Bandits

2019-12-12 · Weihao Kong, Gregory Valiant, Emma Brunskill

We study the problem of estimating the expected reward of the optimal policy in the stochastic disjoint linear bandit setting. We prove that for certain settings it is possible to obtain an accurate estimate of the optim…

Multi-Armed Bandits