中文
相关论文

相关论文: Policy Optimization for Personalized Interventions…

200 篇论文

We apply diffusion strategies to develop a fully-distributed cooperative reinforcement learning algorithm in which agents in a network communicate only with their immediate neighbors to improve predictions about their environment. The…

多智能体系统 · 计算机科学 2014-11-06 Sergio Valcarcel Macua , Jianshu Chen , Santiago Zazo , Ali H. Sayed

Contextual bandits often provide simple and effective personalization in decision making problems, making them popular tools to deliver personalized interventions in mobile health as well as other health applications. However, when bandits…

机器学习 · 计算机科学 2021-07-28 Jiayu Yao , Emma Brunskill , Weiwei Pan , Susan Murphy , Finale Doshi-Velez

The great behavioral heterogeneity observed between individuals with the same psychiatric disorder and even within one individual over time complicates both clinical practice and biomedical research. However, modern technologies are an…

神经元与认知 · 定量生物学 2023-05-25 Michaela Ennis

Information theory is a powerful tool to express principles to drive autonomous systems because it is domain invariant and allows for an intuitive interpretation. This paper studies the use of the predictive information (PI), also called…

机器人学 · 计算机科学 2013-07-19 Georg Martius , Ralf Der , Nihat Ay

Policy iteration is one of the classical frameworks of reinforcement learning, which requires a known initial stabilizing control. However, finding the initial stabilizing control depends on the known system model. To relax this requirement…

系统与控制 · 电气工程与系统科学 2025-03-20 Dongdong Li , Jiuxiang Dong

We propose a new sample-efficient methodology, called Supervised Policy Update (SPU), for deep reinforcement learning. Starting with data generated by the current policy, SPU formulates and solves a constrained optimization problem in the…

机器学习 · 计算机科学 2018-12-27 Quan Vuong , Yiming Zhang , Keith W. Ross

Estimating optimal dynamic policies from offline data is a fundamental problem in dynamic decision making. In the context of causal inference, the problem is known as estimating the optimal dynamic treatment regime. Even though there exists…

计量经济学 · 经济学 2023-12-15 Qizhao Chen , Morgane Austern , Vasilis Syrgkanis

In consequential domains, it is often impossible to compel individuals to take treatment, so that optimal policy rules are merely suggestions in the presence of human non-adherence to treatment recommendations. We study personalized…

机器学习 · 计算机科学 2026-04-24 Angela Zhou

The CDOI outcome measure - a patient-reported outcome (PRO) instrument utilizing direct client feedback - was implemented in a large, real-world behavioral healthcare setting in order to evaluate previous findings from smaller controlled…

Markovian processes have long been used to model stochastic environments. Reinforcement learning has emerged as a framework to solve sequential planning and decision-making problems in such environments. In recent years, attempts were made…

人工智能 · 计算机科学 2014-01-17 Mahdi Milani Fard , Joelle Pineau

Policy learning can be used to extract individualized treatment regimes from observational data in healthcare, civics, e-commerce, and beyond. One big hurdle to policy learning is a commonplace lack of overlap in the data for different…

机器学习 · 统计学 2020-12-04 Nathan Kallus

The computational underpinnings of positive psychotic symptoms have recently received significant attention. Candidate mechanisms include some combination of maladaptive priors and reduced updating of these priors during perception. A…

计算机与社会 · 计算机科学 2021-03-26 David Benrimoh , Ely Sibarium , Andrew Sheldon , Albert Powers

In several medical decision-making problems, such as antibiotic prescription, laboratory testing can provide precise indications for how a patient will respond to different treatment options. This enables us to "fully observe" all potential…

机器学习 · 计算机科学 2020-08-14 Soorajnath Boominathan , Michael Oberst , Helen Zhou , Sanjat Kanjilal , David Sontag

Effective epidemic control is crucial for mitigating the spread of infectious diseases, particularly when pharmaceutical interventions such as vaccines or treatments are limited. Non-pharmaceutical strategies, including mobility…

物理与社会 · 物理学 2025-12-03 Deborah Volpe , Giacomo Orlandi , Mattia Boggio , Carlo Novara , Lorenzo Zino , Giovanna Turvani

Our goal is to compute a policy that guarantees improved return over a baseline policy even when the available MDP model is inaccurate. The inaccurate model may be constructed, for example, by system identification techniques when the true…

最优化与控制 · 数学 2015-06-17 Yinlam Chow , Marek Petrik , Mohammad Ghavamzadeh

National responses to the Covid-19 pandemic varied markedly across countries, from business-as-usual to complete shutdowns. Policies aimed at disrupting the viral transmission cycle and preventing the healthcare system from being…

种群与进化 · 定量生物学 2021-10-27 Arash Saeidpour , Pejman Rohani

Drug overdose deaths, including those due to prescription opioids, represent a critical public health issue in the United States and worldwide. Artificial intelligence (AI) approaches have been developed and deployed to help prescribers…

Social robots offer a promising solution for autonomously guiding patients through physiotherapy exercise sessions, but effective deployment requires advanced decision-making to adapt to patient needs. A key challenge is the scarcity of…

机器人学 · 计算机科学 2025-09-16 Carl Bettosi , Lynne Ballie , Susan Shenkin , Marta Romeo

Self-guided mental health interventions, such as "do-it-yourself" tools to learn and practice coping strategies, show great promise to improve access to mental health care. However, these interventions are often cognitively demanding and…

人机交互 · 计算机科学 2024-04-12 Ashish Sharma , Kevin Rushton , Inna Wanyin Lin , Theresa Nguyen , Tim Althoff

We present \textbf{EGPF} (Equilibrium-Guided Personalization Framework), a mathematically rigorous architecture unifying Bayesian game theory, category theory, information theory, and generative AI for hyper-personalized physician…

计算机科学与博弈论 · 计算机科学 2026-04-09 Suyash Mishra