paper-with-me

홈 › Papers

Hybrid Control Policy for Artificial Pancreas via Ensemble Deep Reinforcement Learning

2023-07-13 · Wenzhou Lv, Tianyu Wu, Luolin Xiong, Liang Wu, Jian Zhou, Yang Tang, Feng Qian

Objective: The artificial pancreas (AP) has shown promising potential in achieving closed-loop glucose control for individuals with type 1 diabetes mellitus (T1DM). However, designing an effective control policy for the AP remains challenging due to the complex physiological processes, delayed insulin response, and inaccurate glucose measurements. While model predictive control (MPC) offers safety and stability through the dynamic model and safety constraints, it lacks individualization and is adversely affected by unannounced meals. Conversely, deep reinforcement learning (DRL) provides personalized and adaptive strategies but faces challenges with distribution shifts and substantial data requirements. Methods: We propose a hybrid control policy for the artificial pancreas (HyCPAP) to address the above challenges. HyCPAP combines an MPC policy with an ensemble DRL policy, leveraging the strengths of both policies while compensating for their respective limitations. To facilitate faster deployment of AP systems in real-world settings, we further incorporate meta-learning techniques into HyCPAP, leveraging previous experience and patient-shared knowledge to enable fast adaptation to new patients with limited available data. Results: We conduct extensive experiments using the FDA-accepted UVA/Padova T1DM simulator across three scenarios. Our approaches achieve the highest percentage of time spent in the desired euglycemic range and the lowest occurrences of hypoglycemia. Conclusion: The results clearly demonstrate the superiority of our methods for closed-loop glucose management in individuals with T1DM. Significance: The study presents novel control policies for AP systems, affirming the great potential of proposed methods for efficient closed-loop glucose control.

📄 PDF Abstract BibTeX arXiv:2307.06501

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningMeta-LearningModel Predictive Controlreinforcement-learning

Similar Papers 제목 키워드 기반

Reinforcement learning for suppression of collective activity in oscillatory ensembles

2019-09-25 · Dmitriy Krylov, Dmitry V. Dylov, Michael Rosenblum

We present a use of modern data-based machine learning approaches to suppress self-sustained collective oscillations typically signaled by ensembles of degenerative neurons in the brain. The proposed hybrid model relies …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Application of Deep Reinforcement Learning to Event-Triggered Control for Networked Artificial Pancreas Systems

2026-04-28 · Junya Ikemoto, Satoshi Maruyama, Kazumune Hashimoto arxiv

This paper proposes a deep reinforcement learning (DRL)-based event-triggered controller design for networked artificial pancreas (AP) systems. Although existing DRL-based AP controllers typically assume periodic control…

Reinforcement Learning

Pancreas segmentation with probabilistic map guided bi-directional recurrent UNet

2019-03-03 · Jun Li, Xiaozhu Lin, Hui Che, Hao Li 외

Pancreas segmentation in medical imaging data is of great significance for clinical pancreas diagnostics and treatment. However, the large population variations in the pancreas shape and volume cause enormous segmentatio…

Pancreas SegmentationSegmentation

Self-Triggered Control in Artificial Pancreas

2024-11-16 · Debayani Ghosh, Sahaj Saxena, Navin Kumar

The management of type 1 diabetes has been revolutionized by the artificial pancreas system (APS), which automates insulin delivery based on continuous glucose monitor (CGM). While conventional closed-loop systems rely o…

Management

MPC-guided Imitation Learning of Neural Network Policies for the Artificial Pancreas

2020-03-03 · Hongkai Chen, Nicola Paoletti, Scott A. Smolka, Shan Lin

Even though model predictive control (MPC) is currently the main algorithm for insulin control in the artificial pancreas (AP), it usually requires complex online optimizations, which are infeasible for resource-constrai…

Bayesian InferenceImitation LearningModel Predictive ControlState Estimation