English
Related papers

Related papers: Risk-sensitive Semi-Markov Decision Problems with …

200 papers

We consider Markov decision processes where the state of the chain is only given at chosen observation times and of a cost. Optimal strategies involve the optimisation of observation times as well as the subsequent action values. We…

Optimization and Control · Mathematics 2025-03-27 Christoph Reisinger , Jonathan Tam

We consider a class of diffusions controlled through the drift and jump size, and driven by a jump L\'evy process and a nondegenerate Wiener process, and we study infinite horizon (ergodic) risk-sensitive control problem for this model. We…

Optimization and Control · Mathematics 2021-03-02 Ari Arapostathis , Anup Biswas

In many sequential decision-making problems we may want to manage risk by minimizing some measure of variability in rewards in addition to maximizing a standard criterion. Variance related risk measures are among the most common…

Machine Learning · Computer Science 2015-03-19 Prashanth L. A. , Mohammad Ghavamzadeh

Traditional solvable optimal control theory predominantly focuses on quadratic costs due to their analytical tractability, yet they often fail to capture critical non-linearities inherent in real-world systems including water, energy,…

Optimization and Control · Mathematics 2025-05-22 Julian Barreiro-Gomez , Tyrone E. Duncan , Bozenna Pasik-Duncan , Hamidou Tembine

We consider robust Markov Decision Processes with Borel state and action spaces, unbounded cost and finite time horizon. Our formulation leads to a Stackelberg game against nature. Under integrability, continuity and compactness assumptions…

Optimization and Control · Mathematics 2025-10-16 Nicole Bäuerle , Alexander Glauner

This work addresses the problem of risk-sensitive control for nonlinear systems with imperfect state observations, extending results for the linear case. In particular, we derive an algorithm that can compute local solutions with…

Optimization and Control · Mathematics 2021-10-22 Bilal Hammoud , Armand Jordana , Ludovic Righetti

In this paper, co-states are used to develop a framework that desensitizes the optimal cost. A general formulation for an optimal control problem with fixed final time is considered. The proposed scheme involves elevating the parameters of…

Optimization and Control · Mathematics 2019-10-02 Venkata Ramana Makkapati , Dipankar Maity , Mehregan Dor , Panagiotis Tsiotras

We use classical tools from calculus of variations to formally derive necessary conditions for a Markov control to be optimal in a standard finite time horizon stochastic control problem. As an example, we solve the well-known Merton…

Optimization and Control · Mathematics 2026-05-27 Matthew Lorig

In many practical sequential decision-making problems, tracking the state of the environment incurs a sensing/communication/computation cost. In these settings, the agent's interaction with its environment includes the additional component…

Machine Learning · Computer Science 2026-04-16 Vansh Kapoor , Jayakrishnan Nair

Safety-critical cyber-physical systems require control strategies whose worst-case performance is robust against adversarial disturbances and modeling uncertainties. In this paper, we present a framework for approximate control and learning…

Optimization and Control · Mathematics 2023-04-04 Aditya Dave , Ioannis Faros , Nishanth Venkatesh , Andreas A. Malikopoulos

This note describes sufficient conditions under which total-cost and average-cost Markov decision processes (MDPs) with general state and action spaces, and with weakly continuous transition probabilities, can be reduced to discounted MDPs.…

Optimization and Control · Mathematics 2017-11-21 Eugene A. Feinberg , Jefferson Huang

Partially observable Markov decision processes (POMDPs) provide an elegant mathematical framework for modeling complex decision and planning problems in stochastic domains in which states of the system are observable only indirectly, via a…

Artificial Intelligence · Computer Science 2011-06-02 M. Hauskrecht

A discrete-time method for solving problems in optimal quantum control is presented. Controlling the time discretized markovian dynamics of a quantum system can be reduced to a Markov-decision process. We demonstrate this method in this…

Quantum Physics · Physics 2012-06-05 Jon R. Grice , David A. Meyer

Economic Model Predictive Control has recently gained popularity due to its ability to directly optimize a given performance criterion, while enforcing constraint satisfaction for nonlinear systems. Recent research has developed both…

Systems and Control · Electrical Eng. & Systems 2022-01-25 Mario Zanon , Sébastien Gros

Markov decision processes (MDPs) are the defacto frame-work for sequential decision making in the presence ofstochastic uncertainty. A classical optimization criterion forMDPs is to maximize the expected discounted-sum pay-off, which…

Artificial Intelligence · Computer Science 2020-02-28 Tomas Brazdil , Krishnendu Chatterjee , Petr Novotny , Jiri Vahala

In this paper we investigate the local risk-minimization approach for a semimartingale financial market where there are restrictions on the available information to agents who can observe at least the asset prices. We characterize the…

Probability · Mathematics 2014-11-20 Claudia Ceci , Katia Colaneri , Alessandra Cretarola

A multiplicative relative value iteration algorithm for solving the dynamic programming equation for the risk-sensitive control problem is studied for discrete time controlled Markov chains with a compact Polish state space, and controlled…

Optimization and Control · Mathematics 2019-12-19 Ari Arapostathis , Vivek S. Borkar

In this article we consider zero and non-zero sum risk-sensitive average criterion games for semi-Markov processes with a finite state space. For the zero-sum case, under suitable assumptions we show that the game has a value. We also…

Optimization and Control · Mathematics 2021-06-10 Arnab Bhabak , Subhamay Saha

We consider an auto-scaling technique in a cloud system where virtual machines hosted on a physical node are turned on and off depending on the queue's occupation (or thresholds), in order to minimise a global cost integrating both energy…

Optimization and Control · Mathematics 2021-07-26 Thomas Tournaire , Hind Castel-Taleb , Emmanuel Hyon

This paper investigates the limit behavior of Markov Decision Processes (MDPs) made of independent particles evolving in a common environment, when the number of particles goes to infinity. In the finite horizon case or with a discounted…

Probability · Mathematics 2009-06-10 Nicolas Gast , Bruno Gaujal