paper-with-me

홈 › Papers

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care

2026-01-30 · Joel Romero-Hernandez, Oscar Camara arxiv

Pain management in intensive care usually involves complex trade-offs, since both inadequate and excessive treatment can compromise patient safety. Prior work on reinforcement learning for sedation and analgesia has explored how to optimize these interventions, but has not considered patient survival or partial observability. To investigate the risks of these design choices, we developed an offline deep reinforcement learning framework that suggests hourly medication doses based on recurrent state representations. Using retrospective data from 47,144 ICU stays in the MIMIC-IV database, we trained and evaluated behavior-regularized actor-critic models that prescribe continuous doses of opioids, propofol, benzodiazepines, and dexmedetomidine according to two goals: reduce pain or jointly reduce pain and 30-day post-discharge mortality. Although the two resulting policies were associated with lower pain, clinician agreement with the pain-only policy was positively correlated with mortality ($ρ$=0.119, p<0.0001), while agreement with the joint policy was negatively correlated ($ρ$=-0.316, p<0.0001). We found that such divergence arose from a different response to high levels of comorbidity. This suggests that valuing post-discharge outcomes could be critical for learning safer treatment policies, even if a short-term goal remains the primary objective.

📄 PDF Abstract BibTeX arXiv:2601.23154

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units

2017-04-20 · Niranjani Prasad, Li-Fang Cheng, Corey Chivers, Michael Draugelis 외

The management of invasive mechanical ventilation, and the regulation of sedation and analgesia during ventilation, constitutes a major part of the care of patients admitted to intensive care units. Both prolonged depend…

Managementreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Estimation of Clinical Workload and Patient Activity using Deep Learning and Optical Flow

2022-02-09 · Thanh Nguyen-Duc, Peter Y Chan, Andrew Tay, David Chen 외

Contactless monitoring using thermal imaging has become increasingly proposed to monitor patient deterioration in hospital, most recently to detect fevers and infections during the COVID-19 pandemic. In this letter, we p…

Motion Estimationobject-detectionObject DetectionOptical Flow Estimation

Pain Detection in Masked Faces during Procedural Sedation

2022-11-12 · Y. Zarghami, S. Mafeld, A. Conway, B. Taati

Pain monitoring is essential to the quality of care for patients undergoing a medical procedure with sedation. An automated mechanism for detecting pain could improve sedation dose titration. Previous studies on facial p…

Medical Procedure

DOSE-I: A Multimodal Biosignal Dataset of Procedural Sedation for Endoscopy -- Technical Report

2026-06-30 · Jakob Garbe, Jan W. Kantelhardt, Katja Seeliger, Thomas Schmid arxiv

In this document, we describe characteristics and technical details of the multimodal biosignal dataset DOSE-I of procedural sedation for endoscopy published on zenodo. The DOSE-I dataset includes 78.5 hours of recording…

Artifact Detection

Safer-Instruct: Aligning Language Models with Automated Preference Data

2023-11-15 · Taiwei Shi, Kai Chen, Jieyu Zhao

Reinforcement learning from human feedback (RLHF) is a vital strategy for enhancing model capability in language models. However, annotating preference data for RLHF is a resource-intensive and creativity-demanding proce…

Diversity