paper-with-me

홈 › Papers

Formally Verifying and Explaining Sepsis Treatment Policies with COOL-MC

2026-02-16 · Dennis Gross arxiv

Safe and interpretable sequential decision-making is critical in healthcare, yet reinforcement learning (RL) policies for sepsis treatment optimization remain opaque and difficult to verify. Standard probabilistic model checkers operate on the full state space, which becomes infeasible for larger MDPs, and cannot explain why a learned policy makes particular decisions. COOL-MC wraps the model checker Storm but adds three key capabilities: it constructs only the reachable state space induced by a trained policy, yielding a smaller discrete-time Markov chain amenable to verification even when full-MDP analysis is intractable; it automatically labels states with clinically meaningful atomic propositions; and it integrates explainability methods with probabilistic computation tree logic (PCTL) queries to reveal which features drive decisions across treatment trajectories. We demonstrate COOL-MC's capabilities on the ICU-Sepsis MDP, a benchmark derived from approximately 17,000 sepsis patient records, which serves as a case study for applying COOL-MC to the formal analysis of sepsis treatment policies. Our analysis establishes hard bounds via full MDP verification, trains a safe RL policy that achieves optimal survival probability, and analyzes its behavior via PCTL verification and explainability on the induced DTMC. This reveals, for instance, that our trained policy relies predominantly on prior dosing history rather than the patient's evolving condition, a weakness that is invisible to standard evaluation but is exposed by COOL-MC's integration of formal verification and explainability. Our results illustrate how COOL-MC could serve as a tool for clinicians to investigate and debug sepsis treatment policies before deployment.

📄 PDF Abstract BibTeX arXiv:2602.14505

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Model-Based Reinforcement Learning for Sepsis Treatment

2018-11-23 · Aniruddh Raghu, Matthieu Komorowski, Sumeetpal Singh

Sepsis is a dangerous condition that is a leading cause of patient mortality. Treating sepsis is highly challenging, because individual patients respond very differently to medical interventions and there is no universal…

modelModel-based Reinforcement Learningreinforcement-learningReinforcement Learning+1

Deep Reinforcement Learning for Sepsis Treatment

2017-11-27 · Aniruddh Raghu, Matthieu Komorowski, Imran Ahmed, Leo Celi 외

Sepsis is a leading cause of mortality in intensive care units and costs hospitals billions annually. Treating a septic patient is highly challenging, because individual patients respond very differently to medical inter…

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement Learning+2

Continuous State-Space Models for Optimal Sepsis Treatment - a Deep Reinforcement Learning Approach

2017-05-23 · Aniruddh Raghu, Matthieu Komorowski, Leo Anthony Celi, Peter Szolovits 외

Sepsis is a leading cause of mortality in intensive care units (ICUs) and costs hospitals billions annually. Treating a septic patient is highly challenging, because individual patients respond very differently to medica…

Decision MakingDeep Reinforcement LearningReinforcement LearningReinforcement Learning (RL)+1

Offline reinforcement learning with uncertainty for treatment strategies in sepsis

2021-07-09 · Ran Liu, Joseph L. Greenstein, James C. Fackler, Jules Bergmann 외

Guideline-based treatment for sepsis and septic shock is difficult because sepsis is a disparate range of life-threatening organ dysfunctions whose pathophysiology is not fully understood. Early intervention in sepsis is…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

How Consistent are Clinicians? Evaluating the Predictability of Sepsis Disease Progression with Dynamics Models

2024-04-10 · Unnseo Park, Venkatesh Sivaraman, Adam Perer

Reinforcement learning (RL) is a promising approach to generate treatment policies for sepsis patients in intensive care. While retrospective evaluation metrics show decreased mortality when these policies are followed, …

DiversityReinforcement Learning (RL)