paper-with-me

홈 › Papers

Estimating Optimal Infinite Horizon Dynamic Treatment Regimes via pT-Learning

2021-10-20 · Wenzhuo Zhou, Ruoqing Zhu, Annie Qu

Recent advances in mobile health (mHealth) technology provide an effective way to monitor individuals' health statuses and deliver just-in-time personalized interventions. However, the practical use of mHealth technology raises unique challenges to existing methodologies on learning an optimal dynamic treatment regime. Many mHealth applications involve decision-making with large numbers of intervention options and under an infinite time horizon setting where the number of decision stages diverges to infinity. In addition, temporary medication shortages may cause optimal treatments to be unavailable, while it is unclear what alternatives can be used. To address these challenges, we propose a Proximal Temporal consistency Learning (pT-Learning) framework to estimate an optimal regime that is adaptively adjusted between deterministic and stochastic sparse policy models. The resulting minimax estimator avoids the double sampling issue in the existing algorithms. It can be further simplified and can easily incorporate off-policy data without mismatched distribution corrections. We study theoretical properties of the sparse policy and establish finite-sample bounds on the excess risk and performance error. The proposed method is provided in our proximalDTR package and is evaluated through extensive simulation studies and the OhioT1DM mHealth dataset.

📄 PDF Abstract BibTeX arXiv:2110.10719

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Constructing Dynamic Treatment Regimes in Infinite-Horizon Settings

2014-06-03 · Ashkan Ertefaie

The application of existing methods for constructing optimal dynamic treatment regimes is limited to cases where investigators are interested in optimizing a utility function over a fixed period of time (finite horizon).…

Nutrition

Estimating Dynamic Treatment Regimes in Mobile Health Using V-learning

2016-11-10 · Daniel J. Luckett, Eric B. Laber, Anna R. Kahkoska, David M. Maahs 외

The vision for precision medicine is to use individual patient characteristics to inform a personalized treatment plan that leads to the best healthcare possible for each patient. Mobile technologies have an important ro…

Decision MakingReinforcement Learning

Inference on Optimal Dynamic Policies via Softmax Approximation

2023-03-08 · Qizhao Chen, Morgane Austern, Vasilis Syrgkanis

Estimating optimal dynamic policies from offline data is a fundamental problem in dynamic decision making. In the context of causal inference, the problem is known as estimating the optimal dynamic treatment regime. Even…

Causal InferenceDecision Makingvalid

Beyond dynamic programming

2023-06-26 · Abhinav Muraleedharan

In this paper, we present Score-life programming, a novel theoretical approach for solving reinforcement learning problems. In contrast with classical dynamic programming-based methods, our method can search over non-sta…

reinforcement-learningReinforcement Learning

Quantile Markov Decision Process

2017-11-15 · Xiaocheng Li, Huaiyang Zhong, Margaret L. Brandeau

The goal of a traditional Markov decision process (MDP) is to maximize expected cumulativereward over a defined horizon (possibly infinite). In many applications, however, a decision maker may beinterested in optimizing …