中文
相关论文

相关论文: Game and Reference: Policy Combination Synthesis f…

200 篇论文

Mathematical models are increasing adopted for setting targets for disease prevention and control. As model-informed policies are implemented, however, the inaccuracies of some forecasts become apparent, for example overprediction of…

We introduce a new preference-based framework for conditional treatment effect estimation and policy learning, built on the Conditional Preference-based Treatment Effect (CPTE). CPTE requires only that outcomes be ranked under a preference…

机器学习 · 统计学 2026-02-04 Dovid Parnas , Mathieu Even , Julie Josse , Uri Shalit

This paper introduces a microscopic approach to model epidemics, which can explicitly consider the consequences of individual's decisions on the spread of the disease. We first formulate a microscopic multi-agent epidemic model where every…

多智能体系统 · 计算机科学 2020-04-28 Changliu Liu

AI agents are being developed to support high stakes decision-making processes from driving cars to prescribing drugs, making it increasingly important for human users to understand their behavior. Policy summarization methods aim to convey…

机器学习 · 计算机科学 2019-06-03 Isaac Lage , Daphna Lifschitz , Finale Doshi-Velez , Ofra Amir

Background: The global spread of the severe acute respiratory syndrome (SARS) epidemic has clearly shown the importance of considering the long-range transportation networks in the understanding of emerging diseases outbreaks. The…

其他定量生物学 · 定量生物学 2008-01-16 Vittoria Colizza , Alain Barrat , Marc Barthelemy , Alessandro Vespignani

Modelling epidemics via classical population-based models suffers from shortcomings that so-called individual-based models are able to overcome, as they are able to take heterogeneity features into account, such as super-spreaders, and…

最优化与控制 · 数学 2022-05-16 C. Courtès , E. Franck , K. Lutz , L. Navoret , Y. Privat

During the COVID-19 pandemic of 2019/2020, authorities have used temporary ad-hoc policy measures, such as lockdowns and mass quarantines, to slow its transmission. However, the consequences of widespread use of these unprecedented measures…

The implementation of public policies is crucial in controlling the spread of COVID-19. However, the effectiveness of different policies can vary across different aspects of epidemic containment. Identifying the most effective policies is…

应用统计 · 统计学 2024-08-27 Zihan Wang

Pandemics involve the high transmission of a disease that impacts global and local health and economic patterns. The impact of a pandemic can be minimized by enforcing certain restrictions on a community. However, while minimizing infection…

人工智能 · 计算机科学 2024-02-13 Ishir Rao

Epidemic models describe the evolution of a communicable disease over time. These models are often modified to include the effects of interventions (control measures) such as vaccination, social distancing, school closings etc. Many such…

统计方法学 · 统计学 2026-01-27 Heejong Bong , Valérie Ventura , Larry Wasserman

The high sample complexity of reinforcement learning challenges its use in practice. A promising approach is to quickly adapt pre-trained policies to new environments. Existing methods for this policy adaptation problem typically rely on…

机器学习 · 计算机科学 2020-06-16 Yuda Song , Aditi Mavalankar , Wen Sun , Sicun Gao

The effective control of the COVID-19 pandemic is one the most challenging issues of nowadays. The design of optimal control policies is perplexed from a variety of social, political, economical and epidemiological factors. Here, based on…

动力系统 · 数学 2023-03-16 Antonis Armaou , Bryce Katch , Lucia Russo , Constantinos Siettos

The recent pandemic emphasized the need to consider the role of human behavior in shaping epidemic dynamics. In particular, it is necessary to extend beyond the classical epidemiological structures to fully capture the interplay between the…

动力系统 · 数学 2025-01-22 Leah LeJeune , Navid Ghaffarzadegan , Lauren Childs , Omar Saucedo

Policy gradient (PG) methods are successful approaches to deal with continuous reinforcement learning (RL) problems. They learn stochastic parametric (hyper)policies by either exploring in the space of actions or in the space of parameters.…

机器学习 · 计算机科学 2024-05-31 Alessandro Montenegro , Marco Mussi , Alberto Maria Metelli , Matteo Papini

Psychological defense mechanisms (PDMs) are unconscious cognitive processes that modulate how individuals perceive and respond to emotional distress. Automatically classifying PDMs from text is clinically valuable but severely hindered by…

计算与语言 · 计算机科学 2026-05-15 Hoang-Thuy-Duong Vu , Quoc-Cuong Pham , Huy-Hieu Pham

The ability to compute reward-optimal policies for given and known finite Markov decision processes (MDPs) underpins a variety of applications across planning, controller synthesis, and verification. However, we often want policies (1) to…

计算机科学中的逻辑 · 计算机科学 2025-11-18 Linus Heck , Filip Macák , Milan Češka , Sebastian Junges

Severe infectious diseases such as the novel coronavirus (COVID-19) pose a huge threat to public health. Stringent control measures, such as school closures and stay-at-home orders, while having significant effects, also bring huge economic…

机器学习 · 计算机科学 2022-03-01 Runzhe Wan , Xinyu Zhang , Rui Song

Epidemiological models can not only be used to forecast the course of a pandemic like COVID-19, but also to propose and design non-pharmaceutical interventions such as school and work closing. In general, the design of optimal policies…

最优化与控制 · 数学 2023-04-06 Jan-Hendrik Niemann , Samuel Uram , Sarah Wolf , Nataša Djurdjevac Conrad , Martin Weiser

Model-based Reinforcement Learning approaches have the promise of being sample efficient. Much of the progress in learning dynamics models in RL has been made by learning models via supervised learning. But traditional model-based…

机器学习 · 计算机科学 2019-06-12 Shagun Sodhani , Anirudh Goyal , Tristan Deleu , Yoshua Bengio , Sergey Levine , Jian Tang

Counterfactual estimation using synthetic controls is one of the most successful recent methodological developments in causal inference. Despite its popularity, the current description only considers time series aligned across units and…

机器学习 · 统计学 2021-02-03 Alexis Bellot , Mihaela van der Schaar