English
Related papers

Related papers: Persistent-Transient Policy Evaluation for Markov …

200 papers

Poisson's equation plays a fundamental role as a tool for performance evaluation and optimization of Markov chains. For continuous-time birth-death chains with possibly unbounded transition and cost rates as addressed herein, when…

Probability · Mathematics 2022-07-28 José Niño-Mora

We consider Markov chains on partially ordered sets that generalize the success-runs and remaining life chains in reliability theory. We find conditions for recurrence and transience and give simple expressions for the invariant…

Probability · Mathematics 2010-04-08 Kyle Siegrist

Optimal designs minimize the number of experimental runs (samples) needed to accurately estimate model parameters, resulting in algorithms that, for instance, efficiently minimize parameter estimate variance. Governed by knowledge of past…

Methodology · Statistics 2023-02-03 Nicholas W. Barendregt , Emily G. Webb , Zachary P. Kilpatrick

In this paper, we study consistent and partially exchangeable sequences of Markov chains on a finite state space. We provide a characterisation of the admissible transition rates via a decomposition into individual and coordinated motion of…

This work deals with tailored reduced order models for bifurcating nonlinear parametric partial differential equations, where multiple coexisting solutions arise for a given parametric instance. Approaches based on proper orthogonal…

Numerical Analysis · Mathematics 2025-05-14 Federico Pichi , Maria Strazzullo

The concepts of probability, statistics and stochastic theory are being successfully used in structural engineering. Markov Chain modelling is a simple stochastic process model that has found its application in both describing stochastic…

Applications · Statistics 2007-08-14 K. Balaji Rao

In this paper, we investigate system theoretic properties of transient average constrained economic model predictive control (MPC) without terminal constraints. We show that the optimal open-loop solution passes by the optimal steady-state…

Systems and Control · Electrical Eng. & Systems 2020-10-21 Mario Rosenfelder , Johannes Köhler , Frank Allgöwer

We consider finite-state Markov chains that can be naturally decomposed into smaller ``projection'' and ``restriction'' chains. Possibly this decomposition will be inductive, in that the restriction chains will be smaller copies of the…

Probability · Mathematics 2007-05-23 Mark Jerrum , Jung-Bae Son , Prasad Tetali , Eric Vigoda

We consider Markov chains that obey the following general non-linear state space model: $\Phi_{k+1} = F(\Phi_k, \alpha(\Phi_k, U_{k+1}))$ where the function $F$ is $C^1$ while $\alpha$ is typically discontinuous and $\{U_k: k \in…

Probability · Mathematics 2019-02-07 Alexandre Chotard , Anne Auger

We study group-averaged Markov chains obtained by augmenting a $\pi$-stationary transition kernel $P$ with a group action on the state space via orbit kernels. Given a group $\mathcal{G}$ with orbits $(\mathcal{O}_i)_{i=1}^k$, we analyse…

Probability · Mathematics 2025-12-16 Michael C. H. Choi , Ryan J. Y. Lim , Youjia Wang

Though reinforcement learning has greatly benefited from the incorporation of neural networks, the inability to verify the correctness of such systems limits their use. Current work in explainable deep learning focuses on explaining only a…

Machine Learning · Computer Science 2019-05-30 Nicholay Topin , Manuela Veloso

In this paper, we study the problem of transient signal analysis. A signal-dependent algorithm is proposed which sequentially identifies the countable sets of decay rates and expansion coefficients present in a given signal. We…

Optimization and Control · Mathematics 2015-09-21 Tarek A. Lahlou , Anuran Makur

Policy Iteration (PI) is a classical family of algorithms to compute an optimal policy for any given Markov Decision Problem (MDP). The basic idea in PI is to begin with some initial policy and to repeatedly update the policy to one from an…

We consider infinite-horizon $\gamma$-discounted Markov Decision Processes, for which it is known that there exists a stationary optimal policy. We consider the algorithm Value Iteration and the sequence of policies $\pi_1,...,\pi_k$ it…

Artificial Intelligence · Computer Science 2012-04-02 Bruno Scherrer

We study continuous-time Markov chains on the non-negative integers under mild regularity conditions (in particular, the set of jump vectors is finite and both forward and backward jumps are possible). Based on the so-called flux balance…

Probability · Mathematics 2024-11-26 Mads Chr Hansen , Carsten Wiuf , Chuang Xu

The classical policy gradient method is the theoretical and conceptual foundation of modern policy-based reinforcement learning (RL) algorithms. Most rigorous analyses of such methods, particularly those establishing convergence guarantees,…

Machine Learning · Computer Science 2026-02-11 Jongmin Lee , Ernest K. Ryu

Ergodic properties and asymptotic stationarity are investigated in this paper for the pseudo-covariance matrix (PCM) of a recursive state estimator which is robust against parametric uncertainties and is based on plant output measurements…

Systems and Control · Computer Science 2016-10-12 Tong Zhou

A discrete-time Markov chain can be transformed into a new Markov chain by looking at its states along iterations of an almost surely finite stopping time. By the optional stopping theorem, any bounded harmonic function with respect to the…

Probability · Mathematics 2022-05-04 Iddo Ben-Ari , Behrang Forghani

We consider evaluating improper priors in a formal Bayes setting according to the consequences of their use. Let $\Phi$ be a class of functions on the parameter space and consider estimating elements of $\Phi$ under quadratic loss. If the…

Statistics Theory · Mathematics 2011-09-07 Brian P. Shea , Galin L. Jones

We develop a general theory for Markov chains whose transition probabilities are the coefficients of descent operators on combinatorial Hopf algebras. These model the breaking-then-recombining of combinational objects. Examples include the…

Combinatorics · Mathematics 2018-08-28 C. Y. Amy Pang