中文
相关论文

相关论文: Learning Individualized Treatment Rules with Estim…

200 篇论文

Randomized experiments are the gold standard for causal inference but face significant challenges in business applications, including limited traffic allocation, the need for heterogeneous treatment effect estimation, and the complexity of…

统计方法学 · 统计学 2025-08-18 Zhenkang Peng , Chengzhang Li , Ying Rong , Renyu Zhang

We study the problem of policy evaluation and learning from batched contextual bandit data when treatments are continuous, going beyond previous work on discrete treatments. Previous work for discrete treatment/action spaces focuses on…

机器学习 · 统计学 2018-02-19 Nathan Kallus , Angela Zhou

Much attention has been devoted recently to the development of machine learning algorithms with the goal of improving treatment policies in healthcare. Reinforcement learning (RL) is a sub-field within machine learning that is concerned…

A treatment regime is a function that maps individual patient information to a recommended treatment, hence explicitly incorporating the heterogeneity in need for treatment across individuals. Patient responses are dichotomous and can be…

机器学习 · 统计学 2016-07-07 Yingfei Wang , Warren Powell

As critically ill patients frequently develop anemia or coagulopathy, transfusion of blood products is a frequent intervention in the Intensive Care Units (ICU). However, inappropriate transfusion decisions made by physicians are often…

机器学习 · 计算机科学 2022-06-30 Yuqing Wang , Yun Zhao , Linda Petzold

Randomized Controlled Trials (RCTs), or A/B testing, have become the gold standard for optimizing various operational policies on online platforms. However, RCTs on these platforms typically cover a limited number of discrete treatment…

计量经济学 · 经济学 2026-02-06 Zhiqi Zhang , Zhiyu Zeng , Ruohan Zhan , Dennis Zhang

With the increasing adoption of electronic health records, there is an increasing interest in developing individualized treatment rules, which recommend treatments according to patients' characteristics, from large observational data.…

统计方法学 · 统计学 2021-05-05 Muxuan Liang , Young-Geun Choi , Yang Ning , Maureen A Smith , Ying-Qi Zhao

Inverse Reinforcement Learning (IRL) -- the problem of learning reward functions from demonstrations of an \emph{expert policy} -- plays a critical role in developing intelligent systems. While widely used in applications, theoretical…

机器学习 · 统计学 2024-02-13 Lei Zhao , Mengdi Wang , Yu Bai

Randomized clinical trials are the gold standard for analyzing treatment effects, but high costs and ethical concerns can limit recruitment, potentially leading to invalid inferences. Incorporating external trial data with similar…

统计方法学 · 统计学 2024-09-09 Yujia Gu , Hanzhong Liu , Wei Ma

Individual treatment effect (ITE) estimation is to evaluate the causal effects of treatment strategies on some important outcomes, which is a crucial problem in healthcare. Most existing ITE estimation methods are designed for centralized…

机器学习 · 计算机科学 2025-03-10 Changchang Yin , Hong-You Chen , Wei-Lun Chao , Ping Zhang

Individualized treatments are crucial for optimal decision making and treatment allocation, specifically in personalized medicine based on the estimation of an individual's dose-response curve across a continuum of treatment levels, e.g.,…

统计方法学 · 统计学 2025-11-20 Max Sampson , Kung-Sik Chan

Using offline observational data for policy evaluation and learning allows decision-makers to evaluate and learn a policy that connects characteristics and interventions. Most existing literature has focused on either discrete treatment…

人工智能 · 计算机科学 2025-01-22 Cheuk Hang Leung , Yiyan Huang , Yijun Li , Qi Wu

Safe reinforcement learning (RL) aims to learn policies that satisfy certain constraints before deploying them to safety-critical applications. Previous primal-dual style approaches suffer from instability issues and lack optimality…

机器学习 · 计算机科学 2022-06-20 Zuxin Liu , Zhepeng Cen , Vladislav Isenbaev , Wei Liu , Zhiwei Steven Wu , Bo Li , Ding Zhao

With the recent advancements of technology in facilitating real-time monitoring and data collection, "just-in-time" interventions can be delivered via mobile devices to achieve both real-time and long-term management and control.…

统计方法学 · 统计学 2023-09-26 Wenzhuo Zhou , Yuhan Li , Ruoqing Zhu

The goal of precision medicine is to provide individualized treatment at each stage of chronic diseases, a concept formalized by Dynamic Treatment Regimes (DTR). These regimes adapt treatment strategies based on decision rules learned from…

统计方法学 · 统计学 2025-06-09 Sophia Yazzourh , Nicolas Savy , Philippe Saint-Pierre , Michael R. Kosorok

Individualized treatment rules tailor treatments to patients based on clinical, demographic, and other characteristics. Estimation of individualized treatment rules requires the identification of individuals who benefit most from the…

统计方法学 · 统计学 2024-06-06 Junwei Shen , Erica E. M. Moodie , Shirin Golchi

Estimating heterogeneous treatment effects is central to data-driven decision-making, yet industrial applications often face a fundamental tension between limited randomized controlled trial (RCT) budgets and abundant but biased…

统计方法学 · 统计学 2026-02-26 Jiacan Gao , Xinyan Su , Mingyuan Ma , Yiyan Huang , Xiao Xu , Xinrui Wan , Tianqi Gu , Enyun Yu , Jiecheng Guo , Zhiheng Zhang

There is increasing interest in allocating treatments based on observed individual characteristics: examples include targeted marketing, individualized credit offers, and heterogeneous pricing. Treatment personalization introduces…

计量经济学 · 经济学 2023-04-06 Evan Munro

A treatment regime is a rule that assigns a treatment to patients based on their covariate information. Recently, estimation of the optimal treatment regime that yields the greatest overall expected clinical outcome of interest has…

统计方法学 · 统计学 2022-03-07 Kevin Gunn , Wenbin Lu , Rui Song

In complex reinforcement learning (RL) problems, policies with similar rewards may have substantially different behaviors. It remains a fundamental challenge to optimize rewards while also discovering as many diverse strategies as possible,…

机器学习 · 计算机科学 2023-10-24 Wei Fu , Weihua Du , Jingwei Li , Sunli Chen , Jingzhao Zhang , Yi Wu