English
Related papers

Related papers: Ergodic-risk Criterion for Stochastically Stabiliz…

200 papers

Markov chains are the de facto finite-state model for stochastic dynamical systems, and Markov decision processes (MDPs) extend Markov chains by incorporating non-deterministic behaviors. Given an MDP and rewards on states, a classical…

Logic in Computer Science · Computer Science 2024-11-13 Krishnendu Chatterjee , Laurent Doyen

Policy optimization algorithms are crucial in many fields but challenging to grasp and implement, often due to complex calculations related to Markov decision processes and varying use of discount and average reward setups. This paper…

Systems and Control · Electrical Eng. & Systems 2025-04-07 Shuang Wu

Stochastic approximation (SA) is an iterative algorithm for finding the fixed point of an operator using noisy samples and widely used in optimization and Reinforcement Learning (RL). The noise in RL exhibits a Markovian structure, and in…

Machine Learning · Computer Science 2025-05-13 Shaan Ul Haque , Sajad Khodadadian , Siva Theja Maguluri

In this article, we study the ergodic risk-sensitive control problem for controlled regime-switching diffusions. Under a blanket stability hypothesis, we solve the associated nonlinear eigenvalue problem for weakly coupled systems and…

Optimization and Control · Mathematics 2022-07-18 Anup Biswas , Somnath Pradhan

The main challenge for adaptive regulation of linear-quadratic systems is the trade-off between identification and control. An adaptive policy needs to address both the estimation of unknown dynamics parameters (exploration), as well as the…

Systems and Control · Computer Science 2019-04-01 Mohamad Kazem Shirani Faradonbeh , Ambuj Tewari , George Michailidis

In this paper we study infinite horizon nonzero-sum stochastic games for controlled discrete-time Markov chains on a Polish state space with risk-sensitive ergodic cost criterion. Under suitable assumptions we show that the associated…

Optimization and Control · Mathematics 2024-08-26 Bivakar Bose , Chandan Pal , Somnath Pradhan , Subhamay Saha

In the simplest sequential decision problem for an ergodic stochastic process X, at each time n a decision u_n is made as a function of past observations X_0,...,X_{n-1}, and a loss l(u_n,X_n) is incurred. In this setting, it is known that…

Probability · Mathematics 2015-02-04 Ramon van Handel

To profit from price oscillations, investors frequently use threshold-type strategies where changes in the portfolio position are triggered by some indicators reaching prescribed levels. In this paper, we investigate threshold-type…

Probability · Mathematics 2022-07-19 Attila Lovas , Miklós Rásonyi

In this paper we consider the problem of obtaining sharp bounds for the performance of temporal difference (TD) methods with linear function approximation for policy evaluation in discounted Markov decision processes. We show that a simple…

Machine Learning · Statistics 2024-06-18 Sergey Samsonov , Daniil Tiapkin , Alexey Naumov , Eric Moulines

Stochastic and soft optimal policies resulting from entropy-regularized Markov decision processes (ER-MDP) are desirable for exploration and imitation learning applications. Motivated by the fact that such policies are sensitive with…

Machine Learning · Computer Science 2022-01-03 Tien Mai , Patrick Jaillet

We consider stationary stochastic dynamical systems evolving on a compact metric space, by perturbing a deterministic dynamics with a random noise, added according to an arbitrary probabilistic distribution. We prove the maximal and…

Dynamical Systems · Mathematics 2018-07-10 Eleonora Catsigeras

Motivated by robotic surveillance applications, this paper studies the novel problem of maximizing the return time entropy of a Markov chain, subject to a graph topology with travel times and stationary distribution. The return time entropy…

Optimization and Control · Mathematics 2018-05-29 Xiaoming Duan , Mishel George , Francesco Bullo

The entropic risk measure is widely used in high-stakes decision-making across economics, management science, finance, and safety-critical control systems because it captures tail risks associated with uncertain losses. However, when data…

Optimization and Control · Mathematics 2026-01-05 Utsav Sadana , Erick Delage , Angelos Georghiou

Two-timescale stochastic approximation (TTSA) is among the most general frameworks for iterative stochastic algorithms. This includes well-known stochastic optimization methods such as SGD variants and those designed for bilevel or minimax…

Machine Learning · Statistics 2024-02-15 Jie Hu , Vishwaraj Doshi , Do Young Eun

In this paper, we concern with the ergodic linear-quadratic closed-loop optimal control problems, in which the state equation is the mean-field stochastic differential equation with periodic coefficients. We first study the asymptotic…

Optimization and Control · Mathematics 2025-05-09 Jiacheng Wu , Qi Zhang

This paper is a survey of various proofs of the so called {\em fundamental theorem of Markov chains}: every ergodic Markov chain has a unique positive stationary distribution and the chain attains this distribution in the limit independent…

Probability · Mathematics 2022-04-05 Somenath Biswas

We consider a discrete time hidden Markov model where the signal is a stationary Markov chain. When conditioned on the observations, the signal is a Markov chain in a random environment under the conditional measure. It is shown that this…

Probability · Mathematics 2009-09-24 Ramon van Handel

Stochastic optimal control problems have a long tradition in applied probability, with the questions addressed being of high relevance in a multitude of fields. Even though theoretical solutions are well understood in many scenarios, their…

Statistics Theory · Mathematics 2024-05-28 Sören Christensen , Claudia Strauch , Lukas Trottner

We consider a stationary regularly varying time series which can be expressedas a function of a geometrically ergodic Markov chain. We obtain practical conditionsfor the weak convergence of the tail array sums and feasible estimators…

Statistics Theory · Mathematics 2018-09-25 Rafal Kulik , Philippe Soulier , Olivier Wintenberger , Rafa Kulik

We study a regulation problem for stochastic systems subject to both continuous fluctuations and rare but significant shocks, modeled as a jump-diffusion with uncertainty in both the drift and the jump intensity. Such settings arise in…

Optimization and Control · Mathematics 2026-05-26 Abel Azze , Bernardo D'Auria , Giorgio Ferrari