English
Related papers

Related papers: Insulin Regimen ML-based control for T2DM patients

200 papers

The mathematical modelling of biological systems has historically followed one of two approaches: comprehensive and minimal. In comprehensive models, the involved biological pathways are modelled independently, then brought together as an…

Dynamical Systems · Mathematics 2021-11-16 Eric Ng , Jaycee Morgan Kaufman , Lennaert van Veen , Yan Fossat

This work proposes a smartphone video-based approach for the estimation of blood glucose in a non-invasive way. Videos using smartphone camera are collected from the tip of the subjects finger and the frames are subsequently converted into…

Signal Processing · Electrical Eng. & Systems 2019-12-02 Tauseef Tasin Chowdhury , Tahmin Mishma , Md. Saeem Osman , Tanzilur Rahman

This paper presents a model-free reinforcement learning (RL) algorithm to synthesize a control policy that maximizes the satisfaction probability of linear temporal logic (LTL) specifications. Due to the consideration of environment and…

Formal Languages and Automata Theory · Computer Science 2022-01-04 Mingyu Cai , Shaoping Xiao , Baoluo Li , Zhiliang Li , Zhen Kan

Modeling the dynamics of probability distributions from time-dependent data samples is a fundamental problem in many fields, including digital health. The goal is to analyze how the distribution of a biomarker, such as glucose, changes over…

Machine Learning · Statistics 2025-09-18 Antonio Álvarez-López , Marcos Matabuena

We present the first model-free Reinforcement Learning (RL) algorithm to synthesise policies for an unknown Markov Decision Process (MDP), such that a linear time property is satisfied. The given temporal property is converted into a Limit…

Machine Learning · Computer Science 2019-02-19 Mohammadhosein Hasanbeig , Alessandro Abate , Daniel Kroening

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

Optimization and Control · Mathematics 2025-12-18 Lingjia Zeng , Manman Li

Optimally sequencing experimental assays in drug discovery is a high-stakes planning problem under severe uncertainty and resource constraints. A primary obstacle for standard reinforcement learning (RL) is the absence of an explicit…

Machine Learning · Computer Science 2026-01-22 Tianchi Chen , Jan Bima , Sean L. Wu , Otto Ritter , Bingjia Yang , Xiang Yu

An accurately identified maximum tolerated dose (MTD) serves as the cornerstone of successful subsequent phases in oncology drug development. Bayesian logistic regression model (BLRM) is a popular and versatile model-based dose-finding…

Methodology · Statistics 2021-05-17 Hongtao Zhang , Alan Y Chiang , Jixian Wang

The global prevalence of diabetes, particularly type 2 diabetes mellitus (T2DM), is rapidly increasing, posing significant health and economic challenges. T2DM not only disrupts blood glucose regulation but also damages vital organs such as…

Machine Learning · Computer Science 2025-06-09 Praveen Kumar , Vincent T. Metzger , Scott A. Malec

Oral Glucose Tolerance Test (OGTT) is one of many way to produce data in the study of the diabetes dynamic. In a recent paper [1.]:\textit{ Estimating insulin sensitivity and $ \beta $-cell function from the oral glucose tolerance test:…

Probability · Mathematics 2024-12-17 Paul Bekima

Patients with type 2 diabetes need to closely monitor blood sugar levels as their routine diabetes self-management. Although many treatment agents aim to tightly control blood sugar, hypoglycemia often stands as an adverse event. In…

Methodology · Statistics 2024-03-15 Yingfa Xie , Haoda Fu , Yuan Huang , Vladimir Pozdnyakov , Jun Yan

A tenet of reinforcement learning is that the agent always observes rewards. However, this is not true in many realistic settings, e.g., a human observer may not always be available to provide rewards, sensors may be limited or…

Machine Learning · Computer Science 2026-03-24 Alireza Kazemipour , Simone Parisi , Matthew E. Taylor , Michael Bowling

Reinforcement Learning (RL) has gained substantial attention across diverse application domains and theoretical investigations. Existing literature on RL theory largely focuses on risk-neutral settings where the decision-maker learns to…

Machine Learning · Computer Science 2024-12-24 Zhengqi Wu , Renyuan Xu

We consider a multi-agent episodic MDP setup where an agent (leader) takes action at each step of the episode followed by another agent (follower). The state evolution and rewards depend on the joint action pair of the leader and the…

Machine Learning · Computer Science 2023-01-10 Arnob Ghosh

We consider non-standard Markov Decision Processes (MDPs) where the target function is not only a simple expectation of the accumulated reward. Instead, we consider rather general functionals of the joint distribution of terminal state and…

Optimization and Control · Mathematics 2025-10-16 Nicole Bäuerle , Tamara Göll , Anna Jaśkiewicz

We introduce a novel approach to large language model (LLM) distillation by formulating it as a constrained reinforcement learning problem. While recent work has begun exploring the integration of task-specific rewards into distillation…

Machine Learning · Computer Science 2025-09-30 Matthieu Zimmer , Xiaotong Ji , Tu Nguyen , Haitham Bou Ammar

In Reinforcement Learning (RL), it is commonly assumed that an immediate reward signal is generated for each action taken by the agent, helping the agent maximize cumulative rewards to obtain the optimal policy. However, in many real-world…

Machine Learning · Computer Science 2024-10-29 Yuting Tang , Xin-Qiang Cai , Yao-Xiang Ding , Qiyu Wu , Guoqing Liu , Masashi Sugiyama

Effective postprandial glucose control is important to glucose management for subjects with diabetes mellitus. In this work, a data-driven meal bolus decision method is proposed without the need of subject-specific glucose management…

Systems and Control · Electrical Eng. & Systems 2021-01-21 Deheng Cai , Wei Liu , Linong Ji , Dawei Shi

It is well known that for any finite state Markov decision process (MDP) there is a memoryless deterministic policy that maximizes the expected reward. For partially observable Markov decision processes (POMDPs), optimal memoryless policies…

Optimization and Control · Mathematics 2016-02-16 Guido Montufar , Keyan Ghazi-Zahedi , Nihat Ay

Within systems biology there is an increasing interest in the stochastic behavior of genetic and biochemical reaction networks. An appropriate stochastic description is provided by the chemical master equation, which represents a continuous…

Biological Physics · Physics 2011-06-23 E. Giampieri , D. Remondini , L. de Oliveira , G. Castellani , P. Lió