English
Related papers

Related papers: On Occupation Time for On-Off Processes with Multi…

200 papers

We study the problem of off-policy policy optimization in Markov decision processes, and develop a novel off-policy policy gradient method. Prior off-policy policy gradient approaches have generally ignored the mismatch between the…

Machine Learning · Computer Science 2019-07-09 Yao Liu , Adith Swaminathan , Alekh Agarwal , Emma Brunskill

Markov switching models are a popular family of models that introduces time-variation in the parameters in the form of their state- or regime-specific values. Importantly, this time-variation is governed by a discrete-valued latent…

Econometrics · Economics 2023-11-13 Yong Song , Tomasz Woźniak

In this paper we study coupled fully non-local equations, where a linear non-local operator jointly acts on the time and space variables. We establish existence and uniqueness of the solution. A maximum principle is proved and used to…

Probability · Mathematics 2025-01-24 Giacomo Ascione , Enrico Scalas , Bruno Toaldo , Lorenzo Torricelli

This paper introduces an analytical formula for the fractional-order conditional moments of nonlinear drift constant elasticity of variance (NLD-CEV) processes under regime switching, governed by continuous-time finite-state irreducible…

Mathematical Finance · Quantitative Finance 2026-02-02 Kittisak Chumpong , Khamron Mekchay , Fukiat Nualsri , Phiraphat Sutthimat

Policy gradient methods are widely adopted reinforcement learning algorithms for tasks with continuous action spaces. These methods succeeded in many application domains, however, because of their notorious sample inefficiency their use…

Machine Learning · Statistics 2024-02-20 Davide Mambelli , Stephan Bongers , Onno Zoeter , Matthijs T. J. Spaan , Frans A. Oliehoek

Several Markovian process calculi have been proposed in the literature, which differ from each other for various aspects. With regard to the action representation, we distinguish between integrated-time Markovian process calculi, in which…

Logic in Computer Science · Computer Science 2010-06-09 Marco Bernardo

This paper is devoted to the study of a stochastic process obtained by random switching between a finite collection of vector fields. Such processes have recently been the focus of much attention in the case where the switching times are…

Probability · Mathematics 2025-10-01 Tobias Hurth , Edouard Strickler

Let $(X_t)_{t \geq 0}$ be a continuous time Markov process on some metric space $M,$ leaving invariant a closed subset $M_0 \subset M,$ called the {\em extinction set}. We give general conditions ensuring either "Stochastic persistence"…

Probability · Mathematics 2023-10-26 Michel Benaim

Model-based methods have recently shown great potential for off-policy evaluation (OPE); offline trajectories induced by behavioral policies are fitted to transitions of Markov decision processes (MDPs), which are used to rollout simulated…

Machine Learning · Computer Science 2023-02-06 Qitong Gao , Ge Gao , Min Chi , Miroslav Pajic

This paper studies a class of optimal multiple stopping problems driven by L\'evy processes. Our model allows for a negative effective discount rate, which arises in a number of financial applications, including stock loans and real…

Mathematical Finance · Quantitative Finance 2016-03-11 Tim Leung , Kazutoshi Yamazaki , Hongzhong Zhang

Reinforcement learning (RL) tasks are typically framed as Markov Decision Processes (MDPs), assuming that decisions are made at fixed time intervals. However, many applications of great importance, including healthcare, do not satisfy this…

We consider a queuing model with the workload evolving between consecutive i.i.d. exponential timers $\{e_q^{(i)}\}_{i=1,2,...}$ according to a spectrally positive L\'{e}vy process $Y(t)$ which is reflected at 0. When the exponential clock…

Probability · Mathematics 2014-04-23 Zbigniew Palmowski , Maria Vlasiou

In this paper we present elementary computations for some Markov modulated counting processes, also called counting processes with regime switching. Regime switching has become an increasingly popular concept in many branches of science. In…

Probability · Mathematics 2023-02-27 Michel Mandjes , Peter Spreij

We consider a finite state discrete time process X. Without loss of generality the finite state space can be identified with the set of unit vectors {e1, e2, . . . , eN} with ei = (0, . . . , 0, 1, 0, . . . , 0)0 2 RN. For a Markov chain…

Probability · Mathematics 2019-05-02 Robert J. Elliott

This work studies the problem of batch off-policy evaluation for Reinforcement Learning in partially observable environments. Off-policy evaluation under partial observability is inherently prone to bias, with risk of arbitrarily large…

Machine Learning · Computer Science 2019-11-26 Guy Tennenholtz , Shie Mannor , Uri Shalit

This paper considers optimization over multiple renewal systems coupled by time average constraints. These systems act asynchronously over variable length frames. For each system, at the beginning of each renewal frame, it chooses an action…

Optimization and Control · Mathematics 2018-05-23 Xiaohan Wei , Michael J. Neely

This paper studies the problem of optimally extracting nonrenewable natural resources. Taking into account the fact that the market values of the main natural resources i.e. oil, natural gas, copper,..., etc, fluctuate randomly following…

General Economics · Economics 2018-07-23 Moustapha Pemy

The distribution of the "mixing time" or the "time to stationarity" in a discrete time irreducible Markov chain, starting in state i, can be defined as the number of trials to reach a state sampled from the stationary distribution of the…

Probability · Mathematics 2014-03-05 Jeffrey J. Hunter

The main topic of these notes are Markov loops, studied in the context of continuous time Markov chains on discrete state spaces. We refer to [1] and [2] for the short "history" of the subject. In contrast with these references, symmetry is…

Probability · Mathematics 2014-02-06 Yinshan Chang , Yves Le Jan

Batch offline data have been shown considerably beneficial for reinforcement learning. Their benefit is further amplified by upsampling with generative models. In this paper, we consider a novel opportunity where interaction with…

Machine Learning · Computer Science 2024-10-04 Shangzhe Li , Xinhua Zhang