中文
相关论文

相关论文: Game and Reference: Policy Combination Synthesis f…

200 篇论文

People's cooperation in adopting protective measures is effective in epidemic control and creates herd immunity as a public good. Similarly, the presence of an epidemic is a driving factor for the formation and improvement of cooperation.…

物理与社会 · 物理学 2025-06-03 Mehran Noori , Nahid Azimi-Tafreshi , Mohammad Salahshour

This paper marries two state-of-the-art controller synthesis methods for partially observable Markov decision processes (POMDPs), a prominent model in sequential decision making under uncertainty. A central issue is to find a POMDP…

计算机科学中的逻辑 · 计算机科学 2023-05-30 Roman Andriushchenko , Alexander Bork , Milan Češka , Sebastian Junges , Joost-Pieter Katoen , Filip Macák

Sepsis is a life-threatening condition defined by end-organ dysfunction due to a dysregulated host response to infection. Although the Surviving Sepsis Campaign has launched and has been releasing sepsis treatment guidelines to unify and…

机器学习 · 计算机科学 2024-11-20 Hyewon Jeong , Siddharth Nayak , Taylor Killian , Sanjat Kanjilal

We provide two methodological insights on \emph{ex ante} policy evaluation for macro models of economic development. First, we show that the problems of parameter instability and lack of behavioral constancy can be overcome by considering…

综合经济学 · 经济学 2019-02-04 Gonzalo Castaeda , Omar A. Guerrero

The opioid epidemic remains one of the most severe public health crises in the United States, yet evaluating policy interventions before implementation is difficult: multiple policies interact within a dynamic system where targeting one…

机器学习 · 计算机科学 2026-02-16 Yijun Ma , Zehong Wang , Weixiang Sun , Zheyuan Zhang , Kaiwen Shi , Nitesh Chawla , Yanfang Ye

In two previous papers, I introduced SuperSpreader (SS) epidemic models, offered some theoretical discussion of prevention issues, and fitted some models to data derived from published accounts of the ongoing MERS epidemic (concluding that…

种群与进化 · 定量生物学 2014-06-24 W. David Wick

In contemporary autonomous driving testing, virtual simulation has become an important approach due to its efficiency and cost effectiveness. However, existing methods usually rely on reinforcement learning to generate risky scenarios,…

机器人学 · 计算机科学 2026-03-24 Chen Xiong , Cheng Wang , Yuhang Liu , Zirui Wu , Ye Tian

We present foundations for using Model Predictive Control (MPC) as a differentiable policy class for reinforcement learning in continuous state and action spaces. This provides one way of leveraging and combining the advantages of…

机器学习 · 计算机科学 2019-10-15 Brandon Amos , Ivan Dario Jimenez Rodriguez , Jacob Sacks , Byron Boots , J. Zico Kolter

Model-based reinforcement learning seeks to simultaneously learn the dynamics of an unknown stochastic environment and synthesise an optimal policy for acting in it. Ensuring the safety and robustness of sequential decisions made through a…

机器学习 · 计算机科学 2023-10-04 Matthew Wicker , Luca Laurenti , Andrea Patane , Nicola Paoletti , Alessandro Abate , Marta Kwiatkowska

Model-based policy optimization is a well-established framework for designing reliable and high-performance controllers across a wide range of control applications. Recently, this approach has been extended to model predictive control…

系统与控制 · 电气工程与系统科学 2026-04-15 Riccardo Zuliani , Efe C. Balta , John Lygeros

Aim of this paper is the description of a new tool to support institutions in the implementation of targeted countermeasures, based on quantitative and multi-scale elements, for the fight and prevention of emergencies, such as the current…

计算机与社会 · 计算机科学 2020-11-12 A. Sebastianelli , F. Mauro , G. Di Cosmo , F. Passarini , M. Carminati , S. L. Ullo

We develop a feedback control method for networked epidemic spreading processes. In contrast to most prior works which consider mean field, open-loop control schemes, the present work develops a novel framework for feedback control of…

最优化与控制 · 数学 2017-03-23 Nicholas J. Watkins , Cameron Nowzari , George J. Pappas

Game dynamics, which describe how agents' strategies evolve over time based on past interactions, can exhibit a variety of undesirable behaviours including convergence to suboptimal equilibria, cycling, and chaos. While central planners can…

系统与控制 · 电气工程与系统科学 2025-11-25 Ilayda Canyakmaz , Iosif Sakos , Wayne Lin , Antonios Varvitsiotis , Georgios Piliouras

In this paper, we leverage the rapid advances in imitation learning, a topic of intense recent focus in the Reinforcement Learning (RL) literature, to develop new sample complexity results and performance guarantees for data-driven Model…

最优化与控制 · 数学 2022-10-18 Kwangjun Ahn , Zakaria Mhammedi , Horia Mania , Zhang-Wei Hong , Ali Jadbabaie

Conjoint analysis, an application of factorial experimental design, is a popular tool in social science research for studying multidimensional preferences. In such political analysis experiments, respondents are often asked to choose…

统计方法学 · 统计学 2025-05-06 Connor T. Jerzak , Priyanshi Chandra , Rishi Hazra

This paper focuses on developing Pareto-optimal estimation and policy learning to identify the most effective treatment that maximizes the total reward from both short-term and long-term effects, which might conflict with each other. For…

机器学习 · 计算机科学 2024-03-13 Yingrong Wang , Anpeng Wu , Haoxuan Li , Weiming Liu , Qiaowei Miao , Ruoxuan Xiong , Fei Wu , Kun Kuang

Movement is how people interact with and affect their environment. For realistic character animation, it is necessary to synthesize such interactions between virtual characters and their surroundings. Despite recent progress in character…

图形学 · 计算机科学 2023-02-03 Mohamed Hassan , Yunrong Guo , Tingwu Wang , Michael Black , Sanja Fidler , Xue Bin Peng

Many of the challenges facing today's reinforcement learning (RL) algorithms, such as robustness, generalization, transfer, and computational efficiency are closely related to compression. Prior work has convincingly argued why minimizing…

机器学习 · 计算机科学 2021-09-08 Benjamin Eysenbach , Ruslan Salakhutdinov , Sergey Levine

Consider a policymaker who wants to decide which intervention to perform in order to change a currently undesirable situation. The policymaker has at her disposal a team of experts, each with their own understanding of the causal…

人工智能 · 计算机科学 2020-05-21 Dalal Alrajeh , Hana Chockler , Joseph Y. Halpern

Probabilities of Causation (PoC) play a fundamental role in decision-making in law, health care and public policy. Nevertheless, their point identification is challenging, requiring strong assumptions, in the absence of which only bounds…

人工智能 · 计算机科学 2023-10-13 Numair Sani , Atalanti A. Mastakouri