中文
相关论文

相关论文: Insulin Regimen ML-based control for T2DM patients

200 篇论文

The mathematical modelling of biological systems has historically followed one of two approaches: comprehensive and minimal. In comprehensive models, the involved biological pathways are modelled independently, then brought together as an…

动力系统 · 数学 2021-11-16 Eric Ng , Jaycee Morgan Kaufman , Lennaert van Veen , Yan Fossat

This work proposes a smartphone video-based approach for the estimation of blood glucose in a non-invasive way. Videos using smartphone camera are collected from the tip of the subjects finger and the frames are subsequently converted into…

信号处理 · 电气工程与系统科学 2019-12-02 Tauseef Tasin Chowdhury , Tahmin Mishma , Md. Saeem Osman , Tanzilur Rahman

This paper presents a model-free reinforcement learning (RL) algorithm to synthesize a control policy that maximizes the satisfaction probability of linear temporal logic (LTL) specifications. Due to the consideration of environment and…

形式语言与自动机理论 · 计算机科学 2022-01-04 Mingyu Cai , Shaoping Xiao , Baoluo Li , Zhiliang Li , Zhen Kan

Modeling the dynamics of probability distributions from time-dependent data samples is a fundamental problem in many fields, including digital health. The goal is to analyze how the distribution of a biomarker, such as glucose, changes over…

机器学习 · 统计学 2025-09-18 Antonio Álvarez-López , Marcos Matabuena

We present the first model-free Reinforcement Learning (RL) algorithm to synthesise policies for an unknown Markov Decision Process (MDP), such that a linear time property is satisfied. The given temporal property is converted into a Limit…

机器学习 · 计算机科学 2019-02-19 Mohammadhosein Hasanbeig , Alessandro Abate , Daniel Kroening

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

最优化与控制 · 数学 2025-12-18 Lingjia Zeng , Manman Li

Optimally sequencing experimental assays in drug discovery is a high-stakes planning problem under severe uncertainty and resource constraints. A primary obstacle for standard reinforcement learning (RL) is the absence of an explicit…

机器学习 · 计算机科学 2026-01-22 Tianchi Chen , Jan Bima , Sean L. Wu , Otto Ritter , Bingjia Yang , Xiang Yu

An accurately identified maximum tolerated dose (MTD) serves as the cornerstone of successful subsequent phases in oncology drug development. Bayesian logistic regression model (BLRM) is a popular and versatile model-based dose-finding…

统计方法学 · 统计学 2021-05-17 Hongtao Zhang , Alan Y Chiang , Jixian Wang

The global prevalence of diabetes, particularly type 2 diabetes mellitus (T2DM), is rapidly increasing, posing significant health and economic challenges. T2DM not only disrupts blood glucose regulation but also damages vital organs such as…

机器学习 · 计算机科学 2025-06-09 Praveen Kumar , Vincent T. Metzger , Scott A. Malec

Oral Glucose Tolerance Test (OGTT) is one of many way to produce data in the study of the diabetes dynamic. In a recent paper [1.]:\textit{ Estimating insulin sensitivity and $ \beta $-cell function from the oral glucose tolerance test:…

概率论 · 数学 2024-12-17 Paul Bekima

Patients with type 2 diabetes need to closely monitor blood sugar levels as their routine diabetes self-management. Although many treatment agents aim to tightly control blood sugar, hypoglycemia often stands as an adverse event. In…

统计方法学 · 统计学 2024-03-15 Yingfa Xie , Haoda Fu , Yuan Huang , Vladimir Pozdnyakov , Jun Yan

A tenet of reinforcement learning is that the agent always observes rewards. However, this is not true in many realistic settings, e.g., a human observer may not always be available to provide rewards, sensors may be limited or…

机器学习 · 计算机科学 2026-03-24 Alireza Kazemipour , Simone Parisi , Matthew E. Taylor , Michael Bowling

Reinforcement Learning (RL) has gained substantial attention across diverse application domains and theoretical investigations. Existing literature on RL theory largely focuses on risk-neutral settings where the decision-maker learns to…

机器学习 · 计算机科学 2024-12-24 Zhengqi Wu , Renyuan Xu

We consider a multi-agent episodic MDP setup where an agent (leader) takes action at each step of the episode followed by another agent (follower). The state evolution and rewards depend on the joint action pair of the leader and the…

机器学习 · 计算机科学 2023-01-10 Arnob Ghosh

We consider non-standard Markov Decision Processes (MDPs) where the target function is not only a simple expectation of the accumulated reward. Instead, we consider rather general functionals of the joint distribution of terminal state and…

最优化与控制 · 数学 2025-10-16 Nicole Bäuerle , Tamara Göll , Anna Jaśkiewicz

We introduce a novel approach to large language model (LLM) distillation by formulating it as a constrained reinforcement learning problem. While recent work has begun exploring the integration of task-specific rewards into distillation…

机器学习 · 计算机科学 2025-09-30 Matthieu Zimmer , Xiaotong Ji , Tu Nguyen , Haitham Bou Ammar

In Reinforcement Learning (RL), it is commonly assumed that an immediate reward signal is generated for each action taken by the agent, helping the agent maximize cumulative rewards to obtain the optimal policy. However, in many real-world…

机器学习 · 计算机科学 2024-10-29 Yuting Tang , Xin-Qiang Cai , Yao-Xiang Ding , Qiyu Wu , Guoqing Liu , Masashi Sugiyama

Effective postprandial glucose control is important to glucose management for subjects with diabetes mellitus. In this work, a data-driven meal bolus decision method is proposed without the need of subject-specific glucose management…

系统与控制 · 电气工程与系统科学 2021-01-21 Deheng Cai , Wei Liu , Linong Ji , Dawei Shi

It is well known that for any finite state Markov decision process (MDP) there is a memoryless deterministic policy that maximizes the expected reward. For partially observable Markov decision processes (POMDPs), optimal memoryless policies…

最优化与控制 · 数学 2016-02-16 Guido Montufar , Keyan Ghazi-Zahedi , Nihat Ay

Within systems biology there is an increasing interest in the stochastic behavior of genetic and biochemical reaction networks. An appropriate stochastic description is provided by the chemical master equation, which represents a continuous…

生物物理 · 物理学 2011-06-23 E. Giampieri , D. Remondini , L. de Oliveira , G. Castellani , P. Lió