English
Related papers

Related papers: Averaging for some simple constrained Markov proce…

200 papers

We study properties of a piecewise deterministic Markov process modeling the changes in concentration of specific antibodies. The evolution of densities of the process is described by a stochastic semigroup. The long-time behaviour of this…

Probability · Mathematics 2020-05-14 Katarzyna Pichór , Ryszard Rudnicki

Markov Decision Processes (Mdps) form a versatile framework used to model a wide range of optimization problems. The Mdp model consists of sets of states, actions, time steps, rewards, and probability transitions. When in a given state and…

The optimization step in many machine learning problems rarely relies on vanilla gradient descent but it is common practice to use momentum-based accelerated methods. Despite these algorithms being widely applied to arbitrary loss…

Disordered Systems and Neural Networks · Physics 2021-10-29 Stefano Sarao Mannelli , Pierfrancesco Urbani

Partially observable Markov decision processes (POMDPs) have recently become popular among many AI researchers because they serve as a natural model for planning under uncertainty. Value iteration is a well-known algorithm for finding…

Artificial Intelligence · Computer Science 2011-06-02 N. L. Zhang , W. Zhang

In this paper, we investigate the concentration properties of cumulative reward in Markov Decision Processes (MDPs), focusing on both asymptotic and non-asymptotic settings. We introduce a unified approach to characterize reward…

Machine Learning · Computer Science 2025-12-04 Borna Sayedana , Peter E. Caines , Aditya Mahajan

We study the performance of a stochastic algorithm based on the power method that adaptively learns the large deviation functions characterizing the fluctuations of additive functionals of Markov processes, used in physics to model…

Statistical Mechanics · Physics 2023-03-30 Francesco Coghi , Hugo Touchette

This paper investigates the random horizon optimal stopping problem for measure-valued piecewise deterministic Markov processes (PDMPs). This is motivated by population dynamics applications, when one wants to monitor some characteristics…

Probability · Mathematics 2018-09-14 Bertrand Cloez , Benoîte de Saporta , Maud Joubaud

In this work, we establish $\mathrm{L}^2$-exponential convergence for a broad class of Piecewise Deterministic Markov Processes recently proposed in the context of Markov Process Monte Carlo methods and covering in particular the Randomized…

Computation · Statistics 2021-08-03 Christophe Andrieu , Alain Durmus , Nikolas Nüsken , Julien Roussel

We provide a framework for speeding up algorithms for time-bounded reachability analysis of continuous-time Markov decision processes. The principle is to find a small, but almost equivalent subsystem of the original system and only analyse…

Systems and Control · Computer Science 2018-07-26 Pranav Ashok , Yuliya Butkova , Holger Hermanns , Jan Křetínský

We consider finite horizon Markov decision processes under performance measures that involve both the mean and the variance of the cumulative reward. We show that either randomized or history-based policies can improve performance. We prove…

Machine Learning · Computer Science 2011-05-02 Shie Mannor , John Tsitsiklis

The dynamics in games involving multiple players, who adaptively learn from their past experience, is not yet well understood. We analyzed a class of stochastic games with Markov strategies in which players choose their actions…

Probability · Mathematics 2018-04-30 Shohei Hidaka

We show that the convergence of finite state space Markov chains to stationarity can often be considerably speeded up by alternating every step of the chain with a deterministic move. Under fairly general conditions, we show that not only…

Probability · Mathematics 2020-08-27 Sourav Chatterjee , Persi Diaconis

We propose a general framework for entropy-regularized average-reward reinforcement learning in Markov decision processes (MDPs). Our approach is based on extending the linear-programming formulation of policy optimization in MDPs to…

Machine Learning · Computer Science 2017-05-23 Gergely Neu , Anders Jonsson , Vicenç Gómez

The problem of constrained Markov decision process is considered. An agent aims to maximize the expected accumulated discounted reward subject to multiple constraints on its costs (the number of constraints is relatively small). A new dual…

Optimization and Control · Mathematics 2022-10-21 Egor Gladin , Maksim Lavrik-Karmazin , Karina Zainullina , Varvara Rudenko , Alexander Gasnikov , Martin Takáč

One of the most widely used methods for solving average cost MDP problems is the value iteration method. This method, however, is often computationally impractical and restricted in size of solvable MDP problems. We propose acceleration…

Optimization and Control · Mathematics 2008-06-03 Oleksandr Shlakhter , Chi-Guhn Lee

Let $Z = (Z_t)_{t\in[0,\infty)}$ be an ergodic Markov process and, for every $n\in\mathbb{N}$, let $Z^n = (Z_{n^2 t})_{t\in[0,\infty)}$ drive a process $X^n$. Classical results show under suitable conditions that the sequence of…

Probability · Mathematics 2018-03-06 Martin Hutzenthaler , Peter Pfaffelhuber , Clemens Printz

Dynamical processes can be classified in various ways as deterministic or stochastic, and continuous or discrete time. All these types can be studied by the path-spaces they generate, and stationary measures on that path-space. Such…

Dynamical Systems · Mathematics 2026-03-19 Suddhasattwa Das

In this paper we present elementary computations for some Markov modulated counting processes, also called counting processes with regime switching. Regime switching has become an increasingly popular concept in many branches of science. In…

Probability · Mathematics 2023-02-27 Michel Mandjes , Peter Spreij

Markov jump processes are continuous-time stochastic processes with a wide range of applications in both natural and social sciences. Despite their widespread use, inference in these models is highly non-trivial and typically proceeds via…

Machine Learning · Computer Science 2023-06-01 Patrick Seifner , Ramses J. Sanchez

We use the abstract method of (local) martingale problems in order to give criteria for convergence of stochastic processes. Extending previous notions, the formulation we use is neither restricted to Markov processes (or semimartingales),…

Probability · Mathematics 2021-08-27 David Criens , Peter Pfaffelhuber , Thorsten Schmidt
‹ Prev 1 4 5 6 7 8 10 Next ›