English
Related papers

Related papers: Persistent-Transient Policy Evaluation for Markov …

200 papers

The basic question in perturbation analysis of Markov chains is: how do small changes in the transition kernels of Markov chains translate to chains in their stationary distributions? Many papers on the subject have shown, roughly, that the…

Probability · Mathematics 2025-08-13 Na Lin , Yuanyuan Liu , Aaron Smith

This paper studies the exponential stability of random matrix products driven by a general (possibly unbounded) state space Markov chain. It is a cornerstone in the analysis of stochastic algorithms in machine learning (e.g. for parameter…

Machine Learning · Statistics 2021-02-02 Alain Durmus , Eric Moulines , Alexey Naumov , Sergey Samsonov , Hoi-To Wai

Consider the problem of approximating the optimal policy of a Markov decision process (MDP) by sampling state transitions. In contrast to existing reinforcement learning methods that are based on successive approximations to the nonlinear…

Machine Learning · Computer Science 2017-10-18 Mengdi Wang

Markov chains are fundamental models for stochastic dynamics, with applications in a wide range of areas such as population dynamics, queueing systems, reinforcement learning, and Monte Carlo methods. Estimating the transition matrix and…

Statistics Theory · Mathematics 2026-01-26 Lasse Leskelä , Maximilien Dreveton

Markov chains are simple yet powerful mathematical structures to model temporally dependent processes. They generally assume stationary data, i.e., fixed transition probabilities between observations/states. However, live, real-world…

Machine Learning · Computer Science 2024-11-27 Kutalmış Coşkun , Borahan Tümer , Bjarne C. Hiller , Martin Becker

The analysis of parametrised systems is a growing field in verification, but the analysis of parametrised probabilistic systems is still in its infancy. This is partly because it is much harder: while there are beautiful cut-off results for…

Logic in Computer Science · Computer Science 2018-04-06 Paul Gainer , Ernst Moritz Hahn , Sven Schewe

We address real-time sampling and estimation of autoregressive Markovian sources in dynamic yet structurally similar multi-hop wireless networks. Each node caches samples from others and communicates over wireless collision channels, aiming…

Machine Learning · Computer Science 2026-01-27 Xingran Chen , Navid NaderiAlizadeh , Alejandro Ribeiro , Shirin Saeedi Bidokhti

Dynamical detection of quantum phases and phase transitions (QPT) in quenched systems with experimentally convenient initial states is a topic of interest from both theoretical and experimental perspectives. Quenched from polarized states,…

Quantum Physics · Physics 2021-06-10 Ceren B. Dağ , Kai Sun

The transition matrix of a Markov chain $(X_k,k\geq 0)$ on a finite or infinite rooted tree is said to be almost upper-directed if, given $X_k$, the node $X_{k+1}$ is either a descendant of $X_k$ or the parent of $X_k$. It is said to be…

Probability · Mathematics 2024-11-12 Luis Fredes , Jean-François Marckert

In this paper, we consider the stability analysis of large-scale distributed networked control systems with random communication delays between linearly interconnected subsystems. The stability analysis is performed in the Markov jump…

Systems and Control · Computer Science 2015-11-13 Kooktae Lee , Raktim Bhattacharya

We study the complexity of central controller synthesis problems for finite-state Markov decision processes, where the objective is to optimize both the expected mean-payoff performance of the system and its stability. We argue that the…

Systems and Control · Computer Science 2013-05-20 Tomáš Brázdil , Krishnendu Chatterjee , Vojtěch Forejt , Antonín Kučera

Howard's Policy Iteration (HPI) is a classic algorithm for solving Markov Decision Problems (MDPs). HPI uses a "greedy" switching rule to update from any non-optimal policy to a dominating one, iterating until an optimal policy is found.…

Artificial Intelligence · Computer Science 2025-05-05 Dibyangshu Mukherjee , Shivaram Kalyanakrishnan

Decisiveness of infinite Markov chains with respect to some (finite or infinite) target set of states is a key property that allows to compute the reachability probability of this set up to an arbitrary precision. Most of the existing works…

Formal Languages and Automata Theory · Computer Science 2023-06-01 Alain Finkel , Serge Haddad , Lina Ye

Estimating the transition dynamics of controlled Markov chains is crucial in fields such as time series analysis, reinforcement learning, and system exploration. Traditional non-parametric density estimation methods often assume independent…

Statistics Theory · Mathematics 2025-05-21 Imon Banerjee , Vinayak Rao , Harsha Honnappa

Parametric Interval Markov Chains (pIMCs) are a specification formalism that extend Markov Chains (MCs) and Interval Markov Chains (IMCs) by taking into account imprecision in the transition probability values: transitions in pIMCs are…

Logic in Computer Science · Computer Science 2017-06-02 Anicet Bart , Benoit Delahaye , Didier Lime , Eric Monfroy , Charlotte Truchet

We propose polynomial-time algorithms to minimise labelled Markov chains whose transition probabilities are not known exactly, have been perturbed, or can only be obtained by sampling. Our algorithms are based on a new notion of an…

Formal Languages and Automata Theory · Computer Science 2021-10-04 Stefan Kiefer , Qiyi Tang

Pseudospectral analysis is fundamental for quantifying the sensitivity and transient behavior of nonnormal matrices, yet its computational cost scales cubically with dimension, rendering it prohibitive for large-scale systems. While…

Numerical Analysis · Mathematics 2026-02-03 Vladimir R. Kostic , Dragana Lj. Cvetkovic , Ljiljana Cvetkovic

We propose a new flexible tensor model for multiple-equation regression that accounts for latent regime changes. The model allows for dynamic coefficients and multi-dimensional covariates that vary across equations. We assume the…

Methodology · Statistics 2024-07-02 Roberto Casarin , Radu Craiu , Qing Wang

This paper considers robust Markov decision processes under parametric transition distributions. We assume that the true transition distribution is uniquely specified by some parametric distribution, and explicitly enforce that the…

Optimization and Control · Mathematics 2022-11-24 Ben Black , Trivikram Dokka , Christopher Kirkbride

This paper introduces a new approach of treating platoon systems using mean-variance control formulation. The underlying system is a controlled switching diffusion in which the random switching process is a continuous-time Markov chain.…

Optimization and Control · Mathematics 2014-01-22 Zhixin Yang , G. Yin , Le Yi Wang , Hongwei Zhang
‹ Prev 1 4 5 6 7 8 10 Next ›