vix.ing · top · new · best · stats · spec

Model Based Reinforcement Learning for Personalized Heparin Dosing

2023/04/19 by Qinyang He, He, Qinyang, Yonatan Mintz +1 · 2 citations
Economics, Econometrics and Finance · Medicine · #49 #Atrial Fibrillation Management and Outcomes #FOS: Biological sciences #FOS: Computer and information sciences #FOS: Mathematics #Health Systems, Economic Evaluations, Quality of Life #Healthcare Technology and Patient Monitoring #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Quantitative Methods (q-bio.QM)

paper · pdf · doi:10.48550/arxiv.2304.10000

openalex publication_date 2023/04/19 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

A key challenge in sequential decision making is optimizing systems safely under partial information. While much of the literature has focused on the cases of either partially known states or partially known dynamics, it is further exacerbated in cases where both states and dynamics are partially known. Computing heparin doses for patients fits this paradigm since the concentration of heparin in the patient cannot be measured directly and the rates at which patients metabolize heparin vary greatly between individuals. While many proposed solutions are model free, they require complex models and have difficulty ensuring safety. However, if some of the structure of the dynamics is known, a model based approach can be leveraged to provide safe policies. In this paper we propose such a framework to address the challenge of optimizing personalized heparin doses. We use a predictive model parameterized individually by patient to predict future therapeutic effects. We then leverage this model using a scenario generation based approach that is capable of ensuring patient safety. We validate our models with numerical experiments by comparing the predictive capabilities of our model against existing machine learning techniques and demonstrating how our dosing algorithm can treat patients in a simulated ICU environment.

Cited by

Related