English
Related papers

Related papers: Persistent-Transient Policy Evaluation for Markov …

200 papers

In this paper a new estimator for the transition density $\pi$ of an homogeneous Markov chain is considered. We introduce an original contrast derived from regression framework and we use a model selection method to estimate $\pi$ under…

Statistics Theory · Mathematics 2015-06-26 Claire Lacour

In the analysis of Markov chains and processes, it is sometimes convenient to replace an unbounded state space with a "truncated" bounded state space. When such a replacement is made, one often wants to know whether the equilibrium behavior…

Probability · Mathematics 2022-06-24 Alex Infanger , Peter W. Glynn

In Markov networks, measurement blackouts with unknown frequency compromise observations such that thermodynamic quantities can no longer be inferred reliably. In particular, the observed currents neither discern equilibrium from…

Statistical Mechanics · Physics 2025-11-19 Alexander M. Maier , Benjamin Häsler , Udo Seifert

We present an algorithm that can efficiently compute a broad class of inferences for discrete-time imprecise Markov chains, a generalised type of Markov chains that allows one to take into account partially specified probabilities and other…

Probability · Mathematics 2019-07-02 Natan T'Joens , Thomas Krak , Jasper De Bock , Gert de Cooman

We exhibit canonical middle-inverse Choice maps within categorical (Free-Variable) Theory of Primitive Recursion as well as in Theory of partial PR maps over the Theory of Primitive Recursion with predicate abstraction. Using these…

Logic · Mathematics 2009-09-08 Michael Pfender

In this work, we study a natural nonparametric estimator of the transition probability matrices of a finite controlled Markov chain. We consider an offline setting with a fixed dataset, collected using a so-called logging policy. We develop…

Machine Learning · Statistics 2026-03-17 Imon Banerjee , Harsha Honnappa , Vinayak Rao

Variable-length Markov chains on finite quivers provide a natural framework for context-dependent stochastic growth under incidence constraints. I study quiver-valued variable-length Markov chains observed through finite boundary windows…

Probability · Mathematics 2026-04-14 Oleg Kiriukhin

A decomposition principle for nonlinear dynamic compartmental systems is introduced in the present paper. This theory is based on the mutually exclusive and exhaustive, analytical and dynamic, novel system and subsystem partitioning…

Systems and Control · Computer Science 2020-11-24 Huseyin Coskun

We derive a finite-sample probabilistic bound on the parameter estimation error of a system identification algorithm for Linear Switched Systems. The algorithm estimates Markov parameters from a single trajectory and applies a variant of…

Machine Learning · Computer Science 2025-05-19 Daniel Racz , Mihaly Petreczky , Balint Daroczy

We develop a practical approach to establish the stability, that is, the recurrence in a given set, of a large class of controlled Markov chains. These processes arise in various areas of applied science and encompass important numerical…

Statistics Theory · Mathematics 2015-02-02 Christophe Andrieu , Vladislav B. Tadić , Matti Vihola

The present paper focuses on the problem of sampling from a given target distribution $\pi$ defined on some general state space. To this end, we introduce a novel class of non-reversible Markov chains, each chain being defined on an…

Computation · Statistics 2023-05-16 Randal Douc , Alain Durmus , Aurélien Enfroy , Jimmy Olsson

A new wave of work on covariance cleaning and nonlinear shrinkage has delivered asymptotically optimal analytical solutions for large covariance matrices. The same framework has been generalized to empirical cross-covariance matrices, whose…

Statistical Finance · Quantitative Finance 2026-01-22 Efstratios Manolakis , Christian Bongiorno , Rosario Nunzio Mantegna

Stochastic discount factor (SDF) processes in dynamic economies admit a permanent-transitory decomposition in which the permanent component characterizes pricing over long investment horizons. This paper introduces an empirical framework to…

Methodology · Statistics 2022-06-06 Timothy Christensen

Policy gradient methods, which have been extensively studied in the last decade, offer an effective and efficient framework for reinforcement learning problems. However, their performances can often be unsatisfactory, suffering from…

Machine Learning · Computer Science 2026-01-27 Shihab Ahmed , El Houcine Bergou , Aritra Dutta , Yue Wang

We develop a model for credit rating migration that accounts for the impact of economic state fluctuations on default probabilities. The joint process for the economic state and the rating is modelled as a time-homogeneous Markov chain.…

Risk Management · Quantitative Finance 2024-03-25 Michael Kalkbrener , Natalie Packham

We investigate whether shallow quantum circuits can accurately reproduce the short-horizon dynamics of discrete-time Markov chains derived from fashion electronic-commerce recommendation links. Transition operators are compiled into…

Quantum Physics · Physics 2025-11-12 Or Peretz , Tai Dinh , Michal Koren

Static regret to a single expert is often the wrong target for strictly online prediction under non-stationarity, where the best expert may switch repeatedly over time. We study Policy-Controlled Generalized Share (PCGS), a general strictly…

Machine Learning · Computer Science 2026-03-31 Hongkai Hu

We consider a discrete time hidden Markov model where the signal is a stationary Markov chain. When conditioned on the observations, the signal is a Markov chain in a random environment under the conditional measure. It is shown that this…

Probability · Mathematics 2009-09-24 Ramon van Handel

In superconducting quantum circuits, decoherence improvements are frequently obtained through process interventions that simultaneously modify surface chemistry, microstructural topology, and device geometry, leaving mechanistic attribution…

Continuous-time Markov decision processes are an important class of models in a wide range of applications, ranging from cyber-physical systems to synthetic biology. A central problem is how to devise a policy to control the system in order…

Systems and Control · Computer Science 2016-06-01 Ezio Bartocci , Luca Bortolussi , Tomǎš Brázdil , Dimitrios Milios , Guido Sanguinetti