English
Related papers

Related papers: Markov cubature rules for polynomial processes

200 papers

In this work we introduce new approximate similarity relations that are shown to be key for policy (or control) synthesis over general Markov decision processes. The models of interest are discrete-time Markov decision processes, endowed…

Systems and Control · Computer Science 2016-06-01 S. Haesaert , S. Esmaeil Zadeh Soudjani , A. Abate

We study the optimization of the expected long-term reward in finite partially observable Markov decision processes over the set of stationary stochastic policies. In the case of deterministic observations, also known as state aggregation,…

Optimization and Control · Mathematics 2022-11-18 Mareike Dressler , Marina Garrote-López , Guido Montúfar , Johannes Müller , Kemal Rose

Markov decision processes are useful models of concurrency optimisation problems, but are often intractable for exhaustive verification methods. Recent work has introduced lightweight approximative techniques that sample directly from…

Logic in Computer Science · Computer Science 2015-03-24 Axel Legay , Sean Sedwards , Louis-Marie Traonouez

This paper describes sufficient conditions for the existence of optimal policies for Partially Observable Markov Decision Processes (POMDPs) with Borel state, observation, and action sets and with the expected total costs. Action sets may…

Optimization and Control · Mathematics 2014-07-02 Eugene A. Feinberg , Pavlo O. Kasyanov , Michael Z. Zgurovsky

We examine a constrained Markov decision process under uncertain transition probabilities, with the uncertainty modeled as deviations from observed transition probabilities. We construct the uncertainty set associated with the deviations…

Optimization and Control · Mathematics 2025-04-15 V Varagapriya

The goal of this paper is to analyze distributional Markov Decision Processes as a class of control problems in which the objective is to learn policies that steer the distribution of a cumulative reward toward a prescribed target law,…

Optimization and Control · Mathematics 2026-02-09 Nicole Bäuerle , Athanasios Vasileiadis

This tutorial describes recently developed general optimality conditions for Markov Decision Processes that have significant applications to inventory control. In particular, these conditions imply the validity of optimality equations and…

Optimization and Control · Mathematics 2016-06-06 Eugene A. Feinberg

An infinite system of point particles placed in $\mathds{R}^d$ is studied. The particles are of two types; they perform random walks in the course of which those of distinct types repel each other. The interaction of this kind induces an…

Probability · Mathematics 2022-07-18 Yuri Kozitsky , Michael Röckner

This paper is dedicated to the investigation of a new numerical method to approximate the optimal stopping problem for a discrete-time continuous state space Markov chain under partial observations. It is based on a two-step discretization…

Optimization and Control · Mathematics 2016-02-16 Benoîte de Saporta , François Dufour , Christophe Nivot

We define the concept of an `open' Markov process, a continuous-time Markov chain equipped with specified boundary states through which probability can flow in and out of the system. External couplings which fix the probabilities of…

Mathematical Physics · Physics 2017-10-03 Blake S. Pollard

Piecewise Deterministic Markov Processes (PDMPs) are studied in a general framework. First, different constructions are proven to be equivalent. Second, we introduce a coupling between two PDMPs following the same differential flow which…

Probability · Mathematics 2021-08-03 Alain Durmus , Arnaud Guillin , Pierre Monmarché

This paper studies continuous-time Markov decision processes under the risk-sensitive average cost criterion. The state space is a finite set, the action space is a Borel space, the cost and transition rates are bounded, and the…

Optimization and Control · Mathematics 2015-12-22 Qingda Wei , Xian Chen

We study the computational complexity of approximating general constrained Markov decision processes. Our primary contribution is the design of a polynomial time $(0,\epsilon)$-additive bicriteria approximation algorithm for finding optimal…

Data Structures and Algorithms · Computer Science 2025-02-12 Jeremy McMahan

A continuous-time Markov process $X$ can be conditioned to be in a given state at a fixed time $T > 0$ using Doob's $h$-transform. This transform requires the typically intractable transition density of $X$. The effect of the $h$-transform…

Probability · Mathematics 2024-09-16 Marc Corstanje , Frank van der Meulen , Moritz Schauer

Lumping a Markov process introduces a coarser level of description that is useful in many contexts and applications. The dynamics on the coarse grained states is often approximated by its Markovian component. In this letter we derive…

Statistical Mechanics · Physics 2012-07-31 David Andrieux

Reinforcement Learning Algorithms are predominantly developed for stationary environments, and the limited literature that considers nonstationary environments often involves specific assumptions about changes that can occur in transition…

Machine Learning · Computer Science 2025-09-25 Ranga Shaarad Ayyagari , Revanth Raj Eega , Ambedkar Dukkipati

We consider a Markov process in continuous time with a finite number of discrete states. The time-dependent probabilities of being in any state of the Markov chain are governed by a set of ordinary differential equations, whose dimension…

Optimization and Control · Mathematics 2014-10-31 Fernando Lopez-Caamal , Tatiana T. Marquez-Lago

For a class of piecewise deterministic Markov processes, the supports of the invariant measures are characterized. This is based on the analysis of controllability properties of an associated deterministic control system. Its invariant…

Dynamical Systems · Mathematics 2018-04-05 Michel Benaïm , Fritz Colonius , Lettau Ralph

Semi-Markov processes are Markovian processes in which the firing time of the transitions is modelled by probabilistic distributions over positive reals interpreted as the probability of firing a transition at a certain moment in time. In…

Formal Languages and Automata Theory · Computer Science 2017-12-04 Mathias Ruggaard Pedersen , Nathanaël Fijalkow , Giorgio Bacci , Kim Guldstrand Larsen , Radu Mardare

In the development of stochastic integration and the theory of semimartingales, Markov processes have been a constant source of inspiration. Despite this historical interweaving, it turned out that semimartingales should be considered the…

Probability · Mathematics 2022-11-29 Sebastian Rickelhoff , Alexander Schnurr