English
Related papers

Related papers: On a Class of Markov Order Estimators Based on PPM…

200 papers

Using the renewal approach we prove exponential inequalities for additive functionals and empirical processes of ergodic Markov chains, thus obtaining counterparts of inequalities for sums of independent random variables. The inequalities…

Probability · Mathematics 2013-10-18 Radosław Adamczak , Witold Bednorz

We study universal decoding over unknown discrete additive channels determined by a finite-state (unifilar) random process. Aiming at low-complexity decoders, we study variants of noise-guessing decoders that use estimators for the…

Information Theory · Computer Science 2025-07-24 Henrique K. Miyamoto , Sheng Yang

It is known that Dobrushin's ergodicity coefficient is one of the effective tools in the investigations of limiting behavior of Markov processes. Several interesting properties of the ergodicity coefficient of a positive mapping defined on…

Functional Analysis · Mathematics 2017-04-26 Nazife Erkurşun Özcan , Farrukh Mukhamedov

We prove explicit error bounds for Markov chain Monte Carlo (MCMC) methods to compute expectations of functions with unbounded stationary variance. We assume that there is a $p\in(1,2)$ so that the functions have finite $L_p$-norm. For…

Statistics Theory · Mathematics 2015-01-27 Daniel Rudolf , Nikolaus Schweizer

We find upper bounds for the probability of underestimation and overestimation errors in penalized likelihood context tree estimation. The bounds are explicit and applies to processes of not necessarily finite memory. We allow for general…

Statistics Theory · Mathematics 2009-03-11 Florencia Leonardi

We generalize to a broader class of decoupled measures a result of Ziv and Merhav on universal estimation of the specific cross (or relative) entropy for a pair of multi-level Markov measures. The result covers pairs of suitably regular…

Information Theory · Computer Science 2026-03-10 Nicholas Barnfield , Raphaël Grondin , Gaia Pozzoli , Renaud Raquépas

Policy optimization algorithms are crucial in many fields but challenging to grasp and implement, often due to complex calculations related to Markov decision processes and varying use of discount and average reward setups. This paper…

Systems and Control · Electrical Eng. & Systems 2025-04-07 Shuang Wu

We resolve the fundamental problem of online decoding with general $n^{th}$ order ergodic Markov chain models. Specifically, we provide deterministic and randomized algorithms whose performance is close to that of the optimal offline…

Machine Learning · Computer Science 2019-05-31 Vikas K. Garg , Tamar Pichkhadze

In this paper, we study planning in stochastic systems, modeled as Markov decision processes (MDPs), with preferences over temporally extended goals. Prior work on temporal planning with preferences assumes that the user preferences form a…

Robotics · Computer Science 2023-03-09 Hazhar Rahmani , Abhishek N. Kulkarni , Jie Fu

Understanding temporal processes and their correlations in time is of paramount importance for the development of near-term technologies that operate under realistic conditions. Capturing the complete multi-time statistics defining a…

Quantum Physics · Physics 2020-04-28 Philip Taranto

We introduce a gradient-based learning method to automatically adapt Markov chain Monte Carlo (MCMC) proposal distributions to intractable targets. We define a maximum entropy regularised objective function, referred to as generalised speed…

Machine Learning · Statistics 2020-01-07 Michalis K. Titsias , Petros Dellaportas

We present a proof of strong consistency of a Ziv-Merhav-type estimator of the cross entropy rate for pairs of hidden-Markov processes. Our proof strategy has two novel aspects: the focus on decoupling properties of the laws and the use of…

Information Theory · Computer Science 2024-10-23 Nicholas Barnfield , Raphaël Grondin , Gaia Pozzoli , Renaud Raquépas

We obtain the sharp estimates on the growth of the uniform norm of orthonormal polynomials for measures satisfying the Steklov condition. This improves the earlier results by Rakhmanov and completely settles a problem by Steklov. The sharp…

Classical Analysis and ODEs · Mathematics 2015-07-28 A. Aptekarev , S. Denisov , D. Tulyakov

We propose a two step strategy for estimating one-dimensional dynamical parameters of a quantum Markov chain, which involves quantum post-processing the output using a coherent quantum absorber and a "pattern counting'' estimator computed…

Quantum Physics · Physics 2025-08-28 Federico Girotti , Alfred Godley , Mădălin Guţă

Continuous-time Markov chains (CTMCs) are popular modeling formalism that constitutes the underlying semantics for real-time probabilistic systems such as queuing networks, stochastic process algebras, and calculi for systems biology. Prism…

Machine Learning · Computer Science 2023-02-20 Giovanni Bacci , Anna Ingólfsdóttir , Kim G. Larsen , Raphaël Reynouard

We consider the reinforcement learning problem for partially observed Markov decision processes (POMDPs) with large or even countably infinite state spaces, where the controller has access to only noisy observations of the underlying…

Machine Learning · Computer Science 2023-07-20 Semih Cayci , Niao He , R. Srikant

The theory of ``Markov-up'' processes is being developed. This is a new class of stochastic processes with ``partial'' markovian features; it could also be called ``one-sided Markov''. Such a behavior may be found in the real world and in…

Probability · Mathematics 2024-07-01 D. O. Kalikaeva

We present a technique for speeding up the convergence of value iteration for partially observable Markov decisions processes (POMDPs). The underlying idea is similar to that behind modified policy iteration for fully observable Markov…

Artificial Intelligence · Computer Science 2013-01-30 Nevin Lianwen Zhang , Stephen S. Lee , Weihong Zhang

An error in the proof of Lemma 2 (ii) in [I. Werner, Math. Proc. Camb. Phil. Soc. 140(2) 333-347 (2006)], which claims the absolute continuity of dynamically defined measures (DDM), is identified. This undermines the assertion of the…

Dynamical Systems · Mathematics 2020-03-26 Ivan Werner

Standard Markov decision process (MDP) and reinforcement learning algorithms optimize the policy with respect to the expected gain. We propose an algorithm which enables to optimize an alternative objective: the probability that the gain is…

Machine Learning · Computer Science 2023-03-06 Vincent Corlay , Jean-Christophe Sibel
‹ Prev 1 8 9 10 Next ›