English
Related papers

Related papers: A Notation for Markov Decision Processes

200 papers

Advances in mobile computing technologies have made it possible to monitor and apply data-driven interventions across complex systems in real time. Markov decision processes (MDPs) are the primary model for sequential decision problems with…

Methodology · Statistics 2018-03-20 Longshaokan Wang , Eric B. Laber , Katie Witkiewitz

Partially observable Markov decision processes (POMDPs) are standard models for dynamic systems with probabilistic and nondeterministic behaviour in uncertain environments. We prove that in POMDPs with long-run average objective, the…

Computer Science and Game Theory · Computer Science 2022-09-29 Krishnendu Chatterjee , Raimundo Saona , Bruno Ziliotto

The Markov assumption (MA) is fundamental to the empirical validity of reinforcement learning. In this paper, we propose a novel Forward-Backward Learning procedure to test MA in sequential decision making. The proposed test does not assume…

Machine Learning · Statistics 2020-02-06 Chengchun Shi , Runzhe Wan , Rui Song , Wenbin Lu , Ling Leng

We present some hardness results on finding the optimal policy for the static formulation of distributionally robust Markov decision processes. We construct problem instances such that when the considered policy class is Markovian and…

Optimization and Control · Mathematics 2026-05-08 Yan Li

Generators of Markov processes on a countable state space can be represented as finite or infinite matrices. One key property is that the off-diagonal entries corresponding to jump rates of the Markov process are non-negative. Here we…

Probability · Mathematics 2020-09-11 Florian Völlering

We develop a model for credit rating migration that accounts for the impact of economic state fluctuations on default probabilities. The joint process for the economic state and the rating is modelled as a time-homogeneous Markov chain.…

Risk Management · Quantitative Finance 2024-03-25 Michael Kalkbrener , Natalie Packham

An analogue of the classical Mecke formula for Poisson point processes is proved for the class of space-time STIT tessellation processes. From this key identity the Markov property of a class of associated random processes is derived. This…

Probability · Mathematics 2017-11-06 Werner Nagel , Linh Ngoc Nguyen , Christoph Thaele , Viola Weiss

Markov decision processes capture sequential decision making under uncertainty, where an agent must choose actions so as to optimize long term reward. The paper studies efficient reasoning mechanisms for Relational Markov Decision Processes…

Artificial Intelligence · Computer Science 2011-11-02 Chenggang Wang , Saket Joshi , Roni Khardon

Matryoshka dolls, the traditional Russian nesting figurines, are known world-wide for each doll's encapsulation of a sequence of smaller dolls. In this paper, we identify a large class of Markov process whose moments are easy to compute by…

Probability · Mathematics 2020-02-26 Andrew Daw , Jamol Pender

We consider almost upper semi-continuous processes defined on a finite Markov chain. The distributions of the functionals associated with the exit from a finite interval are studied. We also consider some modification of these processes.

Probability · Mathematics 2009-09-09 Ievgen Karnaukh

This is a short tutorial of the Case Management Model and Notation (CMMN) version 1.0. It is targeted to readers with knowledge of basic process or workflow modeling, and it covers the complete CMMN notation. A simple complaints process is…

Software Engineering · Computer Science 2016-08-18 Mike A. Marin

Under continuity and recurrence assumptions, we prove that the iteration of successive partial symmetrizations that form a time-homogeneous Markov process, converges to a symmetrization. We cover several settings, including the…

Probability · Mathematics 2018-08-21 Justin Dekeyser , Jean Van Schaftingen

This paper is concentrated on the classification of permutation matrix with the permutation similarity relation, mainly about the canonical form of a permutational similar equivalence class, the cycle matrix decomposition of a permutation…

General Mathematics · Mathematics 2018-07-05 Wenwei Li

This paper discusses algorithms for solving Markov decision processes (MDPs) that have monotone optimal policies. We propose a two-stage alternating convex optimization scheme that can accelerate the search for an optimal policy by…

Systems and Control · Computer Science 2017-04-04 Robert Mattila , Cristian R. Rojas , Vikram Krishnamurthy , Bo Wahlberg

This is lecture notes on the course "Stochastic Processes". In this format, the course was taught in the spring semesters 2017 and 2018 for third-year bachelor students of the Department of Control and Applied Mathematics, School of Applied…

We consider statistical Markov Decision Processes where the decision maker is risk averse against model ambiguity. The latter is given by an unknown parameter which influences the transition law and the cost functions. Risk aversion is…

Optimization and Control · Mathematics 2021-07-21 Nicole Bäuerle , Ulrich Rieder

Labeled continuous-time Markov chains (CTMCs) describe processes subject to random timing and partial observability. In applications such as runtime monitoring, we must incorporate past observations. The timing of these observations matters…

Logic in Computer Science · Computer Science 2024-01-30 Thom Badings , Matthias Volk , Sebastian Junges , Marielle Stoelinga , Nils Jansen

This paper presents a semi-Markov decision process (SMDP) formulation of the satellite task scheduling problem. This formulation can consider multiple operational objectives simultaneously and plan transitions between distinct functional…

Systems and Control · Electrical Eng. & Systems 2019-10-21 Duncan Eddy , Mykel Kochenderfer

We extend the notion of Cantor-Kantorovich distance between Markov chains introduced by (Banse et al., 2023) in the context of Markov Decision Processes (MDPs). The proposed metric is well-defined and can be efficiently approximated given a…

Machine Learning · Computer Science 2024-07-12 Adrien Banse , Venkatraman Renganathan , Raphaël M. Jungers

In this paper, we investigate the concentration properties of cumulative reward in Markov Decision Processes (MDPs), focusing on both asymptotic and non-asymptotic settings. We introduce a unified approach to characterize reward…

Machine Learning · Computer Science 2025-12-04 Borna Sayedana , Peter E. Caines , Aditya Mahajan
‹ Prev 1 8 9 10 Next ›