Offline reinforcement learning for safer blood glucose control in people with type 1 diabetes.

Emerson, Harry; Guy, Matthew; McConville, Ryan

Emerson, Harry; Guy, Matthew; McConville, Ryan.

Afiliação

Emerson H; University of Bristol, 1 Cathedral Square, Bristol, BS1 5TS, United Kingdom. Electronic address: harry.emerson@bristol.ac.uk.
Guy M; University of Bristol, 1 Cathedral Square, Bristol, BS1 5TS, United Kingdom; University Hospital Southampton, Tremona Road, Southampton, SO16 6YD, Hampshire, United Kingdom. Electronic address: matthew.guy@uhs.nhs.uk.
McConville R; University of Bristol, 1 Cathedral Square, Bristol, BS1 5TS, United Kingdom. Electronic address: ryan.mcconville@bristol.ac.uk.

J Biomed Inform ; 142: 104376, 2023 06.

Article em En | MEDLINE | ID: mdl-37149275

RESUMO

The widespread adoption of effective hybrid closed loop systems would represent an important milestone of care for people living with type 1 diabetes (T1D). These devices typically utilise simple control algorithms to select the optimal insulin dose for maintaining blood glucose levels within a healthy range. Online reinforcement learning (RL) has been utilised as a method for further enhancing glucose control in these devices. Previous approaches have been shown to reduce patient risk and improve time spent in the target range when compared to classical control algorithms, but are prone to instability in the learning process, often resulting in the selection of unsafe actions. This work presents an evaluation of offline RL for developing effective dosing policies without the need for potentially dangerous patient interaction during training. This paper examines the utility of BCQ, CQL and TD3-BC in managing the blood glucose of the 30 virtual patients available within the FDA-approved UVA/Padova glucose dynamics simulator. When trained on less than a tenth of the total training samples required by online RL to achieve stable performance, this work shows that offline RL can significantly increase time in the healthy blood glucose range from 61.6±0.3% to 65.3±0.5% when compared to the strongest state-of-art baseline (p<0.001). This is achieved without any associated increase in low blood glucose events. Offline RL is also shown to be able to correct for common and challenging control scenarios such as incorrect bolus dosing, irregular meal timings and compression errors. The code for this work is available at: https://github.com/hemerson1/offline-glucose.

Assuntos

Diabetes Mellitus Tipo 1; Humanos; Diabetes Mellitus Tipo 1/tratamento farmacológico; Hipoglicemiantes/uso terapêutico; Glicemia; Controle Glicêmico; Automonitorização da Glicemia; Insulina; Algoritmos

Palavras-chave

Artificial pancreas; Glucose control; Reinforcement learning; Type 1 diabetes

Texto completo

Imprimir

XML

PubMed Links

Buscar no Google

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Diabetes Mellitus Tipo 1 Limite: Humans Idioma: En Ano de publicação: 2023 Tipo de documento: Article

Texto completo

Imprimir

XML

PubMed Links

Buscar no Google

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Diabetes Mellitus Tipo 1 Limite: Humans Idioma: En Ano de publicação: 2023 Tipo de documento: Article