English
Related papers

Related papers: Persistent-Transient Policy Evaluation for Markov …

200 papers

We consider continuous-space, discrete-time Markov chains on $\mathbb{R}^d$, that admit a finite number $N$ of metastable states. Our main motivation for investigating these processes is to analyse random Poincar\'e maps, which describe…

Probability · Mathematics 2025-08-19 Nils Berglund

We consider Markov-switching regression models, i.e. models for time series regression analyses where the functional relationship between covariates and response is subject to regime switching controlled by an unobservable Markov chain.…

Methodology · Statistics 2015-05-12 Roland Langrock , Thomas Kneib , Richard Glennie , Théo Michelot

Parametric Markov chains have been introduced as a model for families of stochastic systems that rely on the same graph structure, but differ in the concrete transition probabilities. The latter are specified by polynomial constraints for…

Logic in Computer Science · Computer Science 2017-09-08 Lisa Hutschenreiter , Christel Baier , Joachim Klein

Suppose an online platform wants to compare a treatment and control policy, e.g., two different matching algorithms in a ridesharing system, or two different inventory management algorithms in an online retail site. Standard randomized…

Methodology · Statistics 2022-12-27 Peter Glynn , Ramesh Johari , Mohammad Rasouli

We analyse the $\ell^2(\pi)$-convergence rate of irreducible and aperiodic Markov chains with $N$-band transition probability matrix $P$ and with invariant distribution $\pi$. This analysis is heavily based on: first the study of the…

Probability · Mathematics 2016-01-15 Loïc Hervé , James Ledoux

The long-run average payoff per transition (mean payoff) is the main tool for specifying the performance and dependability properties of discrete systems. The problem of constructing a controller (strategy) simultaneously optimizing several…

Artificial Intelligence · Computer Science 2024-12-19 David Klaška , Antonín Kučera , Vojtěch Kůr , Vít Musil , Vojtěch Řehák

The purpose of this paper is to study the time average behavior of Markov chains with transition probabilities being kernels of completely continuous operators, and therefore to provide a sufficient condition for a class of Markov chains…

Probability · Mathematics 2018-11-16 Shizhou Xu

We investigate absorption, i.e., almost sure convergence to an absorbing state, in time-varying (non-homogeneous) discrete-time Markov chains with finite state space. We consider systems that can switch among a finite set of transition…

Systems and Control · Electrical Eng. & Systems 2020-08-18 Yasin Yazicioglu

We derive some key extremal features for $k$th order Markov chains that can be used to understand how the process moves between an extreme state and the body of the process. The chains are studied given that there is an exceedance of a…

Statistics Theory · Mathematics 2023-01-27 Ioannis Papastathopoulos , Adrian Casey , Jonathan A. Tawn

This work is devoted to the almost sure stabilization of adaptive control systems that involve an unknown Markov chain. The control system displays continuous dynamics represented by differential equations and discrete events given by a…

Probability · Mathematics 2008-07-10 Bernard Bercu , Francois Dufour , G. George Yin

We propose a novel method to directly learn a stochastic transition operator whose repeated application provides generated samples. Traditional undirected graphical models approach this problem indirectly by learning a Markov chain model…

Machine Learning · Statistics 2017-11-08 Anirudh Goyal , Nan Rosemary Ke , Surya Ganguli , Yoshua Bengio

We are interested in the analysis of very large continuous-time Markov chains (CTMCs) with many distinct rates. Such models arise naturally in the context of reliability analysis, e.g., of computer network performability analysis, of power…

Logic in Computer Science · Computer Science 2015-07-24 Ernst Moritz Hahn , Holger Hermanns , Ralf Wimmer , Bernd Becker

Markov control algorithms that perform smooth, non-greedy updates of the policy have been shown to be very general and versatile, with policy gradient and Expectation Maximisation algorithms being particularly popular. For these algorithms,…

Systems and Control · Computer Science 2012-02-20 Thomas Furmston , David Barber

We investigate the parameter recovery of Markov-switching ordinary differential processes from discrete observations, where the differential equations are nonlinear additive models. This framework has been widely applied in biological…

Methodology · Statistics 2025-01-03 Katherine Tsai , Mladen Kolar , Sanmi Koyejo

Transformers process tokens in parallel but are temporally shallow: at position $t$, each layer attends to key-value pairs computed based on the previous layer, yielding a depth capped by the number of layers. Recurrent models offer…

Machine Learning · Computer Science 2026-04-24 Costin-Andrei Oncescu , Depen Morwani , Samy Jelassi , Alexandru Meterez , Mujin Kwun , Sham Kakade

In order to model risk aversion in reinforcement learning, an emerging line of research adapts familiar algorithms to optimize coherent risk functionals, a class that includes conditional value-at-risk (CVaR). Because optimizing the…

Machine Learning · Computer Science 2021-03-09 Audrey Huang , Liu Leqi , Zachary C. Lipton , Kamyar Azizzadenesheli

This paper uses the generator approach of Stein's method to analyze the gap between steady-state distributions of Markov chains and diffusion processes. Until now, the standard way to invoke Stein's method for this problem was to use the…

Probability · Mathematics 2022-02-15 Anton Braverman

This paper discusses the stability analysis of linear parameter varying systems with a parameter-dependent delay where the parameters are assumed to be stochastic piecewise constants under spontaneous Poissonian jumps. Based on stochastic…

Systems and Control · Electrical Eng. & Systems 2021-02-10 Muhammad Zakwan

Traditional partial differential equations with constant coefficients often struggle to capture abrupt changes in real-world phenomena, leading to the development of variable coefficient PDEs and Markovian switching models. Recently,…

Machine Learning · Statistics 2024-09-02 Yi Zhang , Zhikun Zhang , Xiangjun Wang

We review criteria for comparing the efficiency of Markov chain Monte Carlo (MCMC) methods with respect to the asymptotic variance of estimates of expectations of functions of state, and show how such criteria can justify ways of combining…

Probability · Mathematics 2025-02-19 Radford M. Neal , Jeffrey S. Rosenthal
‹ Prev 1 3 4 5 6 7 10 Next ›