English
Related papers

Related papers: On a Class of Markov Order Estimators Based on PPM…

200 papers

Entropy regularized Markov decision processes have been widely used in reinforcement learning. This paper is concerned with the primal-dual formulation of the entropy regularized problems. Standard first-order methods suffer from slow…

Optimization and Control · Mathematics 2023-06-13 Haoya Li , Hsiang-fu Yu , Lexing Ying , Inderjit Dhillon

This paper considers the speed of convergence (mixing) of a finite Markov kernel $P$ with respect to the Kullback-Leibler divergence (entropy). Given a Markov kernel one defines either a discrete-time Markov chain (with the $n$-step…

Probability · Mathematics 2024-09-13 Pietro Caputo , Zongchen Chen , Yuzhou Gu , Yury Polyanskiy

The asymptotic behavior for entropy numbers of general Fourier multiplier operators of multiple series with respect to an abstract complete orthonormal system $\{\phi_{\textbf{m}}\}_{\textbf{m}\in \mathbb{N}^d_0}$ on a probability space and…

Functional Analysis · Mathematics 2021-07-12 Sergio Andrés Córdoba Pareja , Jéssica Milaré , Sergio A. Tozoni

Online (also called "recursive" or "adaptive") estimation of fixed model parameters in hidden Markov models is a topic of much interest in times series modelling. In this work, we propose an online parameter estimation algorithm that…

Computation · Statistics 2011-02-16 Olivier Cappé

In queuing theory and related problems, it is very important to know the numerical characteristics of an investigated system - both in stationary and non-stationary modes. In some cases, such characteristics can be calculated, but this is…

Probability · Mathematics 2021-11-25 Galina Zverkina , Mais Farkhadov

Monte Carlo methods -- such as Markov chain Monte Carlo (MCMC) and piecewise deterministic Markov process (PDMP) samplers -- provide asymptotically exact estimators of expectations under a target distribution. There is growing interest in…

Computation · Statistics 2024-09-09 Adrien Corenflos , Matthew Sutton , Nicolas Chopin

Markov automata combine non-determinism, probabilistic branching, and exponentially distributed delays. This compositional variant of continuous-time Markov decision processes is used in reliability engineering, performance evaluation and…

Logic in Computer Science · Computer Science 2017-05-11 Tim Quatmann , Sebastian Junges , Joost-Pieter Katoen

Given access to a single long trajectory generated by an unknown irreducible Markov chain $M$, we simulate an $\alpha$-lazy version of $M$ which is ergodic. This enables us to generalize recent results on estimation and identity testing…

Machine Learning · Statistics 2021-11-02 Sela Fried , Geoffrey Wolfer

We comment on some conceptual and and technical problems related to computational mechanics, point out some errors in several papers, and straighten out some wrong priority claims. We present explicitly the correct algorithm for…

Data Analysis, Statistics and Probability · Physics 2018-04-09 Peter Grassberger

Markov decision processes (MDPs) with rewards are a widespread and well-studied model for systems that make both probabilistic and nondeterministic choices. A fundamental result about MDPs is that their minimal and maximal expected rewards…

Logic in Computer Science · Computer Science 2024-11-26 Kevin Batz , Benjamin Lucien Kaminski , Christoph Matheja , Tobias Winkler

Euclidean Markov decision processes are a powerful tool for modeling control problems under uncertainty over continuous domains. Finite state imprecise, Markov decision processes can be used to approximate the behavior of these infinite…

Artificial Intelligence · Computer Science 2020-06-29 Manfred Jaeger , Giorgio Bacci , Giovanni Bacci , Kim Guldstrand Larsen , Peter Gjøl Jensen

We investigate the problem of synthesizing optimal control policies for Markov decision processes (MDPs) with both qualitative and quantitative objectives. Specifically, our goal is to achieve a given linear temporal logic (LTL) task with…

Systems and Control · Electrical Eng. & Systems 2025-04-08 Yu Chen , Shaoyuan Li , Xiang Yin

Kolmogorov complexity and algorithmic probability are defined only up to an additive resp. multiplicative constant, since their actual values depend on the choice of the universal reference computer. In this paper, we analyze a natural…

Information Theory · Computer Science 2010-03-29 Markus Mueller

This paper surveys various results about Markov chains on general (non-countable) state spaces. It begins with an introduction to Markov chain Monte Carlo (MCMC) algorithms, which provide the motivation and context for the theory which…

Probability · Mathematics 2009-09-29 Gareth O. Roberts , Jeffrey S. Rosenthal

We prove several results concerning classifications, based on successive observations $(X_1,..., X_n)$ of an unknown stationary and ergodic process, for membership in a given class of processes, such as the class of all finite order Markov…

Probability · Mathematics 2008-06-19 Gusztav Morvai , Benjamin Weiss

We proposed a learning algorithm for nonparametric estimation and on-line prediction for general stationary ergodic sources. We prepare histograms each of which estimates the probability as a finite distribution, and mixture them with…

Information Theory · Computer Science 2010-06-29 Joe Suzuki

We introduce Markov Neural Processes (MNPs), a new class of Stochastic Processes (SPs) which are constructed by stacking sequences of neural parameterised Markov transition operators in function space. We prove that these Markov transition…

Machine Learning · Statistics 2023-05-26 Jin Xu , Emilien Dupont , Kaspar Märtens , Tom Rainforth , Yee Whye Teh

We give a short overview of recent results on a specific class of Markov process: the Piecewise Deterministic Markov Processes (PDMPs). We first recall the definition of these processes and give some general results. On more specific cases…

Statistics Theory · Mathematics 2013-09-25 Romain Azaïs , Jean-Baptiste Bardet , Alexandre Genadot , Nathalie Krell , Pierre-André Zitt

We investigate the statistical complexity of estimating the parameters of a discrete-state Markov chain kernel from a single long sequence of state observations. In the finite case, we characterize (modulo logarithmic factors) the minimax…

Machine Learning · Statistics 2020-08-14 Geoffrey Wolfer , Aryeh Kontorovich

We present an algorithm that can efficiently compute a broad class of inferences for discrete-time imprecise Markov chains, a generalised type of Markov chains that allows one to take into account partially specified probabilities and other…

Probability · Mathematics 2019-07-02 Natan T'Joens , Thomas Krak , Jasper De Bock , Gert de Cooman