中文
相关论文

相关论文: Decision-Dependent Distributionally Robust Markov …

200 篇论文

Interval Markov decision processes (IMDPs) generalise classical MDPs by having interval-valued transition probabilities. They provide a powerful modelling tool for probabilistic systems with an additional variation or uncertainty that…

系统与控制 · 计算机科学 2017-07-07 Ernst Moritz Hahn , Vahid Hashemi , Holger Hermanns , Morteza Lahijanian , Andrea Turrini

In this paper, a stochastic SEQIR epidemic model with Markovian regime-switching is proposed and investigated. The governmental policy and implement efficiency are concerned by a generalized incidence function of the susceptible class. We…

概率论 · 数学 2024-02-27 Hongjie Fan , Kai Wang , Yanling Zhu

As the society becomes more dependent on the presence of electricity, the resilience of the power systems gains more importance. This paper develops a decision support method for distribution system operators to restore electricity after an…

系统与控制 · 电气工程与系统科学 2019-11-12 Onur Yigit Arpali , Ugur Can Yilmaz , Ebru Aydin Gol , Burcu Guldur Erkal , Murat Gol

Diffusion policies are a powerful paradigm for robotic control, but fine-tuning them with human preferences is fundamentally challenged by the multi-step structure of the denoising process. To overcome this, we introduce a Unified Markov…

机器人学 · 计算机科学 2026-02-03 Amitesh Vatsa , Zhixian Xie , Wanxin Jin

Probabilistic prediction of stochastic dynamical systems (SDSs) aims to accurately predict the conditional probability distributions of future states. However, accurate probabilistic predictions tightly hinge on accurate distributional…

最优化与控制 · 数学 2026-04-21 Tao Xu , Jianping He

We propose a new, nonparametric approach to learning and representing transition dynamics in Markov decision processes (MDPs), which can be combined easily with dynamic programming methods for policy optimisation and value estimation. This…

机器学习 · 计算机科学 2012-06-22 Steffen Grunewalder , Guy Lever , Luca Baldassarre , Massi Pontil , Arthur Gretton

Optimal control in non-stationary Markov decision processes (MDP) is a challenging problem. The aim in such a control problem is to maximize the long-term discounted reward when the transition dynamics or the reward function can change over…

应用统计 · 统计学 2017-03-03 Taposh Banerjee , Miao Liu , Jonathan P. How

Networks of contacts capable of spreading infectious diseases are often observed to be highly heterogeneous, with the majority of individuals having fewer contacts than the mean, and a significant minority having relatively very many…

物理与社会 · 物理学 2016-12-21 César Parra-Rojas , Thomas House , Alan J. McKane

Public health organizations face the problem of dispensing treatments (i.e., vaccines, antibiotics, and others) to groups of affected populations through "points-of-dispensing" (PODs) during emergency situations, typically in the presence…

最优化与控制 · 数学 2021-05-04 Yijia Wang , Daniel R. Jiang

Solving Markov Decision Processes (MDPs) remains a central challenge in sequential decision-making, especially when dealing with large state spaces and long-term optimization criteria. A key step in Bellman dynamic programming algorithms is…

最优化与控制 · 数学 2025-08-04 Youssef Ait El Mahjoub , Jean-Michel Fourneau , Salma Alouah

In robust Markov decision processes (MDPs), the uncertainty in the transition kernel is addressed by finding a policy that optimizes the worst-case performance over an uncertainty set of MDPs. While much of the literature has focused on…

机器学习 · 计算机科学 2023-03-02 Yue Wang , Alvaro Velasquez , George Atia , Ashley Prater-Bennette , Shaofeng Zou

Recent years have seen a large amount of interest in epidemics on networks as a way of representing the complex structure of contacts capable of spreading infections through the modern human population. The configuration model is a popular…

种群与进化 · 定量生物学 2017-01-23 Frank Ball , Thomas House

We propose a dynamical model for describing the spread of epidemics. This model is an extension of the SIQR (susceptible-infected-quarantined-recovered) and SIRP (susceptible-infected-recovered-pathogen) models used earlier to describe…

物理与社会 · 物理学 2022-12-08 S. P. Lukyanets , I. S. Gandzha , O. V. Kliushnichenko

The goal of a traditional Markov decision process (MDP) is to maximize expected cumulative reward over a defined horizon (possibly infinite). In many applications, however, a decision maker may be interested in optimizing a specific…

人工智能 · 计算机科学 2025-10-16 Xiaocheng Li , Huaiyang Zhong , Margaret L. Brandeau

We study multistage distributionally robust mixed-integer programs under endogenous uncertainty, where the probability distribution of stage-wise uncertainty depends on the decisions made in previous stages. We first consider two ambiguity…

最优化与控制 · 数学 2020-09-28 Xian Yu , Siqian Shen

We develop a novel data-driven robust model predictive control (DDRMPC) approach for automatic control of irrigation systems. The fundamental idea is to integrate both mechanistic models, which describe dynamics in soil moisture variations,…

系统与控制 · 计算机科学 2020-06-16 Chao Shang , Wei-Han Chen , Abraham Duncan Stroock , Fengqi You

Markov Decision Processes (MDPs) have been used to formulate many decision-making problems in science and engineering. The objective is to synthesize the best decision (action selection) policies to maximize expected rewards (minimize…

最优化与控制 · 数学 2015-07-08 Mahmoud El Chamie , Behcet Acikmese

Sequential decisions in volatile, high-stakes settings require more than maximizing expected return; they require principled uncertainty management. This paper presents the Uncertainty-Aware Markov Decision Process (UAMDP), a unified…

机器学习 · 计算机科学 2025-12-19 Michal Koren , Or Peretz , Tai Dinh , Philip S. Yu

In this paper we address the class of Sequential Decision Making (SDM) problems that are characterized by time-varying parameters. These parameter dynamics are either pre-specified or manipulable. At any given time instant the decision…

最优化与控制 · 数学 2022-01-26 Amber Srivastava , S. M. Salapaka

In this paper we consider the problem of computing an $\epsilon$-optimal policy of a discounted Markov Decision Process (DMDP) provided we can only access its transition function through a generative sampling model that given any…

最优化与控制 · 数学 2019-06-07 Aaron Sidford , Mengdi Wang , Xian Wu , Lin F. Yang , Yinyu Ye