English
Related papers

Related papers: Decision-Dependent Distributionally Robust Markov …

200 papers

Interval Markov decision processes (IMDPs) generalise classical MDPs by having interval-valued transition probabilities. They provide a powerful modelling tool for probabilistic systems with an additional variation or uncertainty that…

Systems and Control · Computer Science 2017-07-07 Ernst Moritz Hahn , Vahid Hashemi , Holger Hermanns , Morteza Lahijanian , Andrea Turrini

In this paper, a stochastic SEQIR epidemic model with Markovian regime-switching is proposed and investigated. The governmental policy and implement efficiency are concerned by a generalized incidence function of the susceptible class. We…

Probability · Mathematics 2024-02-27 Hongjie Fan , Kai Wang , Yanling Zhu

As the society becomes more dependent on the presence of electricity, the resilience of the power systems gains more importance. This paper develops a decision support method for distribution system operators to restore electricity after an…

Systems and Control · Electrical Eng. & Systems 2019-11-12 Onur Yigit Arpali , Ugur Can Yilmaz , Ebru Aydin Gol , Burcu Guldur Erkal , Murat Gol

Diffusion policies are a powerful paradigm for robotic control, but fine-tuning them with human preferences is fundamentally challenged by the multi-step structure of the denoising process. To overcome this, we introduce a Unified Markov…

Robotics · Computer Science 2026-02-03 Amitesh Vatsa , Zhixian Xie , Wanxin Jin

Probabilistic prediction of stochastic dynamical systems (SDSs) aims to accurately predict the conditional probability distributions of future states. However, accurate probabilistic predictions tightly hinge on accurate distributional…

Optimization and Control · Mathematics 2026-04-21 Tao Xu , Jianping He

We propose a new, nonparametric approach to learning and representing transition dynamics in Markov decision processes (MDPs), which can be combined easily with dynamic programming methods for policy optimisation and value estimation. This…

Machine Learning · Computer Science 2012-06-22 Steffen Grunewalder , Guy Lever , Luca Baldassarre , Massi Pontil , Arthur Gretton

Optimal control in non-stationary Markov decision processes (MDP) is a challenging problem. The aim in such a control problem is to maximize the long-term discounted reward when the transition dynamics or the reward function can change over…

Applications · Statistics 2017-03-03 Taposh Banerjee , Miao Liu , Jonathan P. How

Networks of contacts capable of spreading infectious diseases are often observed to be highly heterogeneous, with the majority of individuals having fewer contacts than the mean, and a significant minority having relatively very many…

Physics and Society · Physics 2016-12-21 César Parra-Rojas , Thomas House , Alan J. McKane

Public health organizations face the problem of dispensing treatments (i.e., vaccines, antibiotics, and others) to groups of affected populations through "points-of-dispensing" (PODs) during emergency situations, typically in the presence…

Optimization and Control · Mathematics 2021-05-04 Yijia Wang , Daniel R. Jiang

Solving Markov Decision Processes (MDPs) remains a central challenge in sequential decision-making, especially when dealing with large state spaces and long-term optimization criteria. A key step in Bellman dynamic programming algorithms is…

Optimization and Control · Mathematics 2025-08-04 Youssef Ait El Mahjoub , Jean-Michel Fourneau , Salma Alouah

In robust Markov decision processes (MDPs), the uncertainty in the transition kernel is addressed by finding a policy that optimizes the worst-case performance over an uncertainty set of MDPs. While much of the literature has focused on…

Machine Learning · Computer Science 2023-03-02 Yue Wang , Alvaro Velasquez , George Atia , Ashley Prater-Bennette , Shaofeng Zou

Recent years have seen a large amount of interest in epidemics on networks as a way of representing the complex structure of contacts capable of spreading infections through the modern human population. The configuration model is a popular…

Populations and Evolution · Quantitative Biology 2017-01-23 Frank Ball , Thomas House

We propose a dynamical model for describing the spread of epidemics. This model is an extension of the SIQR (susceptible-infected-quarantined-recovered) and SIRP (susceptible-infected-recovered-pathogen) models used earlier to describe…

Physics and Society · Physics 2022-12-08 S. P. Lukyanets , I. S. Gandzha , O. V. Kliushnichenko

The goal of a traditional Markov decision process (MDP) is to maximize expected cumulative reward over a defined horizon (possibly infinite). In many applications, however, a decision maker may be interested in optimizing a specific…

Artificial Intelligence · Computer Science 2025-10-16 Xiaocheng Li , Huaiyang Zhong , Margaret L. Brandeau

We study multistage distributionally robust mixed-integer programs under endogenous uncertainty, where the probability distribution of stage-wise uncertainty depends on the decisions made in previous stages. We first consider two ambiguity…

Optimization and Control · Mathematics 2020-09-28 Xian Yu , Siqian Shen

We develop a novel data-driven robust model predictive control (DDRMPC) approach for automatic control of irrigation systems. The fundamental idea is to integrate both mechanistic models, which describe dynamics in soil moisture variations,…

Systems and Control · Computer Science 2020-06-16 Chao Shang , Wei-Han Chen , Abraham Duncan Stroock , Fengqi You

Markov Decision Processes (MDPs) have been used to formulate many decision-making problems in science and engineering. The objective is to synthesize the best decision (action selection) policies to maximize expected rewards (minimize…

Optimization and Control · Mathematics 2015-07-08 Mahmoud El Chamie , Behcet Acikmese

Sequential decisions in volatile, high-stakes settings require more than maximizing expected return; they require principled uncertainty management. This paper presents the Uncertainty-Aware Markov Decision Process (UAMDP), a unified…

Machine Learning · Computer Science 2025-12-19 Michal Koren , Or Peretz , Tai Dinh , Philip S. Yu

In this paper we address the class of Sequential Decision Making (SDM) problems that are characterized by time-varying parameters. These parameter dynamics are either pre-specified or manipulable. At any given time instant the decision…

Optimization and Control · Mathematics 2022-01-26 Amber Srivastava , S. M. Salapaka

In this paper we consider the problem of computing an $\epsilon$-optimal policy of a discounted Markov Decision Process (DMDP) provided we can only access its transition function through a generative sampling model that given any…

Optimization and Control · Mathematics 2019-06-07 Aaron Sidford , Mengdi Wang , Xian Wu , Lin F. Yang , Yinyu Ye