paper-with-me

홈 › Papers

Robust Policies For Proactive ICU Transfers

2020-02-14 · Julien Grand-Clement, Carri W. Chan, Vineet Goyal, Gabriel Escobar

Patients whose transfer to the Intensive Care Unit (ICU) is unplanned are prone to higher mortality rates than those who were admitted directly to the ICU. Recent advances in machine learning to predict patient deterioration have introduced the possibility of \emph{proactive transfer} from the ward to the ICU. In this work, we study the problem of finding \emph{robust} patient transfer policies which account for uncertainty in statistical estimates due to data limitations when optimizing to improve overall patient care. We propose a Markov Decision Process model to capture the evolution of patient health, where the states represent a measure of patient severity. Under fairly general assumptions, we show that an optimal transfer policy has a threshold structure, i.e., that it transfers all patients above a certain severity level to the ICU (subject to available capacity). As model parameters are typically determined based on statistical estimations from real-world data, they are inherently subject to misspecification and estimation errors. We account for this parameter uncertainty by deriving a robust policy that optimizes the worst-case reward across all plausible values of the model parameters. We show that the robust policy also has a threshold structure under fairly general assumptions. Moreover, it is more aggressive in transferring patients than the optimal nominal policy, which does not take into account parameter uncertainty. We present computational experiments using a dataset of hospitalizations at 21 KNPC hospitals, and present empirical evidence of the sensitivity of various hospital metrics (mortality, length-of-stay, average ICU occupancy) to small changes in the parameters. Our work provides useful insights into the impact of parameter uncertainty on deriving simple policies for proactive ICU transfer that have strong empirical performance and theoretical guarantees.

📄 PDF Abstract BibTeX arXiv:2002.06247

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Simulation-Free Hierarchical Latent Policy Planning for Proactive Dialogues

2024-12-19 · Tao He, Lizi Liao, Yixin Cao, Yuanxing Liu 외

Recent advancements in proactive dialogues have garnered significant attention, particularly for more complex objectives (e.g. emotion support and persuasion). Unlike traditional task-oriented dialogues, proactive dialog…

Hierarchical Reinforcement LearningReinforcement Learning (RL)User Simulation

PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy Optimization

2026-03-20 · Renhong Huang, Ning Tang, Jiarong Xu, Yuxuan Cao 외 arxiv

Social platforms serve as central hubs for information exchange, where user behaviors and platform interventions jointly shape opinions. However, intervention policies like recommendation and content filtering, can unint…

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

2026-06-03 · Kaustav Kundu, Ritvik Shrivastava, Maxim Arap, Nanshu Wang 외 arxiv

We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding \textit{when} to interrupt, and \textit{how} to coach. However, progress …

Development of a Trust-Aware User Simulator for Statistical Proactive Dialog Modeling in Human-AI Teams

2023-04-24 · Matthias Kraus, Ron Riekenbrauck, Wolfgang Minker

The concept of a Human-AI team has gained increasing attention in recent years. For effective collaboration between humans and AI teammates, proactivity is crucial for close coordination and effective communication. Howe…

Open-Ended Question Answering

Communication Policy Evolution for Proactive LLM Agents

2026-06-12 · Xinbei Ma, Jiyang Qiu, Yao Yao, Zheng Wu 외 arxiv

LLM agents have rapidly evolved into autonomous systems, yet a persistent information gap remains between users and agents: communication is costly, while users' identical preferences further limit information exchange. …