English
Related papers

Related papers: Offline Estimation of Controlled Markov Chains: Mi…

200 papers

This paper extends the standard chaining technique to prove excess risk upper bounds for empirical risk minimization with random design settings even if the magnitude of the noise and the estimates is unbounded. The bound applies to many…

Machine Learning · Statistics 2016-09-08 Gábor Balázs , András György , Csaba Szepesvári

Nonparametric identification and maximum likelihood estimation for finite-state hidden Markov models are investigated. We obtain identification of the parameters as well as the order of the Markov chain if the transition probability…

Statistics Theory · Mathematics 2015-10-01 Grigory Alexandrovich , Hajo Holzmann , Anna Leister

The formal verification and controller synthesis for Markov decision processes that evolve over uncountable state spaces are computationally hard and thus generally rely on the use of approximations. In this work, we consider the…

Systems and Control · Computer Science 2018-11-28 Sofie Haesaert , Sadegh Soudjani , Alessandro Abate

In this work, we consider an inhomogeneous (discrete time) Markov chain and are interested in its long time behavior. We provide sufficient conditions to ensure that some of its asymptotic properties can be related to the ones of a…

Probability · Mathematics 2017-11-09 Michel Benaïm , Florian Bouguet , Bertrand Cloez

We obtain the posterior distribution of a random process conditioned on observing the empirical frequencies of a finite sample path. We find under a rather broad assumption on the "dependence structure" of the process, {\em c.f.}…

Probability · Mathematics 2022-03-02 Wenqing Hu , Hong Qian

We study stochastic optimization algorithms for constrained nonconvex stochastic optimization problems with Markovian data. In particular, we focus on the case when the transition kernel of the Markov chain is state-dependent. Such…

Optimization and Control · Mathematics 2022-11-10 Abhishek Roy , Krishnakumar Balasubramanian , Saeed Ghadimi

In this paper, we show how a simulated Markov decision process (MDP) built by the so-called \emph{baseline} policies, can be used to compute a different policy, namely the \emph{simulated optimal} policy, for which the performance of this…

Optimization and Control · Mathematics 2014-10-13 Yinlam Chow , Mohammad Ghavamzadeh

In order to give quantitative estimates for approximating the ergodic limit, we investigate probabilistic limit behaviors of time-averaging estimators of numerical discretizations for a class of time-homogeneous Markov processes, by…

Probability · Mathematics 2023-10-13 Chuchu Chen , Tonghe Dang , Jialin Hong , Guoting Song

We propose a new approach for estimating the finite dimensional transition matrix of a Markov chain using a large number of independent sample paths observed at random times. The sample paths may be observed as few as two times, and the…

Methodology · Statistics 2025-05-20 Daphne Aurouet , Valentin Patilea

This work focuses on optimal harvesting-renewing for a stochastic population. A mixed regular-singular control formulation with a state constraint and regime-switching is introduced. The decision-makers either harvest or renew with finite…

Optimization and Control · Mathematics 2022-11-07 K. Q. Tran , L. T. N. Bich , George Yin

We consider the problem of sampling a multimodal distribution with a Markov chain given a small number of samples from the stationary measure. Although mixing can be arbitrarily slow, we show that if the Markov chain has a $k$th order…

Machine Learning · Computer Science 2024-11-15 Frederic Koehler , Holden Lee , Thuy-Duong Vuong

The first motivation of this paper is to study stationarity and ergodic properties for a general class of time series models defined conditional on an exogenous covariates process. The dynamic of these models is given by an autoregressive…

Statistics Theory · Mathematics 2020-07-16 Paul Doukhan , Michael H. Neumann , Lionel Truquet

This paper introduces ergodic-risk criteria, which capture long-term cumulative risks associated with controlled Markov chains through probabilistic limit theorems--in contrast to existing methods that require assumptions of either finite…

Optimization and Control · Mathematics 2025-12-03 Shahriar Talebi , Na Li

We study data-driven learning of robust stochastic control for infinite-horizon systems with potentially continuous state and action spaces. In many managerial settings--supply chains, finance, manufacturing, services, and dynamic…

Machine Learning · Statistics 2025-11-18 Shengbo Wang , Jason Meng , Nian Si , Jose Blanchet , Zhengyuan Zhou

We study risk-sensitive control of continuous time Markov chains taking values in discrete state space. We study both finite and infinite horizon problems. In the finite horizon problem we characterise the value function via HJB equation…

Optimization and Control · Mathematics 2014-09-16 Mrinal K. Ghosh , Subhamay Saha

In this paper, we study a notion of local stationarity for discrete time Markov chains which is useful for applications in statistics. In the spirit of some locally stationary processes introduced in the literature, we consider triangular…

Statistics Theory · Mathematics 2016-10-06 Lionel Truquet

This paper is concerned with ergodic properties of inhomogeneous Markov processes. Since the transition probabilities depend on initial times, the existing methods to obtain invariant measures for homogeneous Markov processes are not…

Probability · Mathematics 2025-01-24 Zhenxin Liu , Di Lu

We show that large-scale typicality of Markov sample paths implies that the likelihood ratio statistic satisfies a law of iterated logarithm uniformly to the same scale. As a consequence, the penalized likelihood Markov order estimator is…

Probability · Mathematics 2011-08-31 Ramon van Handel

This paper considers maximum likelihood (ML) estimation in a large class of models with hidden Markov regimes. We investigate consistency of the ML estimator and local asymptotic normality for the models under general conditions which allow…

Statistics Theory · Mathematics 2021-12-07 Demian Pouzo , Zacharias Psaradakis , Martin Sola

Applications of stochastic models often involve the evaluation of steady-state performance, which requires solving a set of balance equations. In most cases of interest, the number of equations is infinite or even uncountable. As a result,…

Optimization and Control · Mathematics 2022-04-08 Shukai Li , Sanjay Mehrotra
‹ Prev 1 8 9 10 Next ›