English
Related papers

Related papers: Policy Learning for Optimal Individualized Dose In…

200 papers

For effective decision support in scenarios with conflicting objectives, sets of potentially optimal solutions can be presented to the decision maker. We explore both what policies these sets should contain and how such sets can be computed…

Artificial Intelligence · Computer Science 2023-07-19 Willem Röpke , Conor F. Hayes , Patrick Mannion , Enda Howley , Ann Nowé , Diederik M. Roijers

We study the convergence rates of policy iteration (PI) for nonconvex viscous Hamilton--Jacobi equations using a discrete space-time scheme, where both space and time variables are discretized. We analyze the case with an uncontrolled…

Numerical Analysis · Mathematics 2025-03-05 Xiaoqin Guo , Hung Vinh Tran , Yuming Paul Zhang

In this paper, we initiate a systematic investigation of differentially private algorithms for convex empirical risk minimization. Various instantiations of this problem have been studied before. We provide new algorithms and matching lower…

Machine Learning · Computer Science 2014-10-21 Raef Bassily , Adam Smith , Abhradeep Thakurta

There is tremendous interest in precision medicine as a means to improve patient outcomes by tailoring treatment to individual characteristics. An individualized treatment rule formalizes precision medicine as a map from patient information…

Machine Learning · Statistics 2020-05-28 Daniel J. Luckett , Eric B. Laber , Michael R. Kosorok

This work investigates the challenge of ensuring safety guarantees in the presence of uncontrollable agents, whose behaviors are stochastic and depend on both their own and the system's states. We present a neural model predictive control…

Systems and Control · Electrical Eng. & Systems 2026-04-21 Shuqi Wang , Mingyang Feng , Yu Chen , Yue Gao , Xiang Yin

We study episodic reinforcement learning (RL) in non-stationary linear kernel Markov decision processes (MDPs). In this setting, both the reward function and the transition kernel are linear with respect to the given feature maps and are…

Machine Learning · Computer Science 2024-12-24 Han Zhong , Zhongren Chen , Zhuoran Yang , Zhaoran Wang , Csaba Szepesvári

Treatment effects of stochastic policy shifts quantify differences in outcomes across counterfactual scenarios with varying treatment distributions. Stochastic policy shifts may be of interest in settings where it is unrealistic or…

Methodology · Statistics 2026-03-31 Michael Jetsupphasuk , Chenwei Fang , Didong Li , Michael G. Hudgens

Decisions in public health are almost always made in the context of uncertainty. Policy makers are responsible for making important decisions, faced with the daunting task of choosing from amongst many possible options. This task is called…

Artificial Intelligence · Computer Science 2020-05-19 Atiye Alaeddini , Daniel Klein

In this work, we propose a new local optimization method to solve a class of nonconvex semidefinite programming (SDP) problems. The basic idea is to approximate the feasible set of the nonconvex SDP problem by inner positive semidefinite…

Optimization and Control · Mathematics 2012-02-27 Quoc Tran Dinh , Wim Michiels , Moritz Diehl

In medicine, treatments often influence multiple, interdependent outcomes, such as primary endpoints, complications, adverse events, or other secondary endpoints. Hence, to make optimal treatment decisions, clinicians are interested in…

Machine Learning · Computer Science 2025-06-03 Yuchen Ma , Jonas Schweisthal , Hengrui Zhang , Stefan Feuerriegel

Policy Iteration (PI) is a widely used family of algorithms to compute optimal policies for Markov Decision Problems (MDPs). We derive upper bounds on the running time of PI on Deterministic MDPs (DMDPs): the class of MDPs in which every…

Discrete Mathematics · Computer Science 2023-10-10 Ritesh Goenka , Eashan Gupta , Sushil Khyalia , Pratyush Agarwal , Mulinti Shaik Wajid , Shivaram Kalyanakrishnan

An optimal dynamic treatment regime (DTR) is a sequence of decision rules aimed at providing the best course of treatments individualized to patients. While conventional DTR estimation uses longitudinal data, such data can also be…

Methodology · Statistics 2025-02-06 Larry Dong , Eleanor Pullenayegum , Rodolphe Thiébaut , Olli Saarela

In this study we performed a feasibility investigation on implementing a fast and accurate dose calculation based on a deep learning technique. A two dimensional (2D) fluence map was first converted into a three dimensional (3D) volume…

Medical Physics · Physics 2021-02-03 Jiawei Fan , Lei Xing , Peng Dong , Jiazhou Wang , Weigang Hu , Yong Yang

We consider a distributionally robust Partially Observable Markov Decision Process (DR-POMDP), where the distribution of the transition-observation probabilities is unknown at the beginning of each decision period, but their realizations…

Optimization and Control · Mathematics 2020-12-09 Hideaki Nakao , Ruiwei Jiang , Siqian Shen

To promote precision medicine, individualized treatment regimes (ITRs) are crucial for optimizing the expected clinical outcome based on patient-specific characteristics. However, existing ITR research has primarily focused on scenarios…

Methodology · Statistics 2024-02-20 Chang Wang , Lu Wang

With the advancement of treatment modalities in radiation therapy for cancer patients, outcomes have improved, but at the cost of increased treatment plan complexity and planning time. The accurate prediction of dose distributions would…

Medical Physics · Physics 2018-12-03 Dan Nguyen , Troy Long , Xun Jia , Weiguo Lu , Xuejun Gu , Zohaib Iqbal , Steve Jiang

A fundamental principle of clinical medicine is that a treatment should only be administered to those patients who would benefit from it. Treatment strategies that assign treatment to patients as a function of their individual…

Applications · Statistics 2025-06-13 Nicholas Williams , Kara Rudolph , Iván Díaz

This paper focuses on the problem of modeling and estimating interaction effects between covariates and a continuous treatment variable on an outcome, using a single-index regression approach. The primary motivation is to estimate an…

Methodology · Statistics 2021-02-02 Hyung Park , Eva Petkova , Thaddeus Tarpey , R. Todd Ogden

Identification of optimal dose combinations in early phase dose-finding trials is challenging, due to the trade-off between precisely estimating the many parameters required to flexibly model the possibly non-monotonic dose-response…

Methodology · Statistics 2024-02-13 James Willard , Shirin Golchi , Erica E. M. Moodie , Bruno Boulanger , Bradley P. Carlin

This paper considers optimal control of dynamical systems which are represented by nonlinear stochastic differential equations. It is well-known that the optimal control policy for this problem can be obtained as a function of a value…

Robotics · Computer Science 2014-05-30 Oktay Arslan , Evangelos Theodorou , Panagiotis Tsiotras