English
Related papers

Related papers: Finite Time Analysis of Linear Two-timescale Stoch…

200 papers

In reinforcement learning (RL) , one of the key components is policy evaluation, which aims to estimate the value function (i.e., expected long-term accumulated reward) of a policy. With a good policy evaluation method, the RL algorithms…

Machine Learning · Computer Science 2018-09-25 Yue Wang , Wei Chen , Yuting Liu , Zhi-Ming Ma , Tie-Yan Liu

In this paper, we extend the results of Elliott and Yang \cite{elliott3} and discuss the control of a stochastic process for which the driving noise is provided by a martingale associated with a semi-Markov Chain. An existence and a…

Probability · Mathematics 2025-12-23 Robert J. Elliott , Zhe Yang

This work analyzes the stochastic approximation algorithm with non-decaying gains as applied in time-varying problems. The setting is to minimize a sequence of scalar-valued loss functions $f_k(\cdot)$ at sampling times $\tau_k$ or to…

Optimization and Control · Mathematics 2020-03-18 Jingyi Zhu

We consider the discrete-time filtering problem in scenarios where the observation noise is degenerate or low. More precisely, one is given access to a discrete time observation sequence which at any time $k$ depends only on the state of an…

Computation · Statistics 2025-11-17 Abylay Zhumekenov , Alexandros Beskos , Dan Crisan , Ajay Jasra , Nikolas Kantas

In this contribution, we provide convergence rates for a finite volume scheme of a stochastic non-linear parabolic equation with multiplicative Lipschitz noise and homogeneous Neumann boundary conditions. More precisely, we give an error…

Numerical Analysis · Mathematics 2025-12-22 Kavin Rajasekaran , Niklas Sapountzoglou

We propose a new abstract formalism for probabilistic timed systems, Parametric Interval Probabilistic Timed Automata, based on an extension of Parametric Timed Automata and Interval Markov Chains. In this context, we consider the…

Formal Languages and Automata Theory · Computer Science 2019-06-13 Étienne André , Benoît Delahaye , Paulin Fournier

This paper studies a class of random nonlinear systems with time-varying delay, in which the $r$-order moment ($r\geq1$) of the random disturbance is finite. Firstly, some general conditions are proposed to guarantee the existence and…

Optimization and Control · Mathematics 2018-06-22 Yao Liqiang , Zhang Weihai

This paper considers the smooth bilevel optimization in which the lower-level problem is strongly convex and the upper-level problem is possibly nonconvex. We focus on the stochastic setting where the algorithm can access the unbiased…

Machine Learning · Computer Science 2025-12-16 Zhuanghua Liu , Luo Luo

Motivated by problems arising in decentralized control problems and non-cooperative Nash games, we consider a class of strongly monotone Cartesian variational inequality (VI) problems, where the mappings either contain expectations or their…

Optimization and Control · Mathematics 2013-01-10 Farzad Yousefian , Angelia Nedić , Uday V. Shanbhag

We analyse the asymptotic properties of a continuous-time, two-timescale stochastic approximation algorithm designed for stochastic bilevel optimisation problems in continuous-time models. We obtain the weak convergence rate of this…

Optimization and Control · Mathematics 2022-07-08 Louis Sharrock

This paper considers estimating the parameters in a regime-switching stochastic differential equation(SDE) driven by Normal Inverse Gaussian(NIG) noise. The model under consideration incorporates a continuous-time finite state Markov chain…

Computation · Statistics 2024-12-10 Yuzhong Cheng , Hiroki Masuda

We consider a two-state model of non-Markovian stochastic resonance (SR) within the framework of the theory of renewal processes. Residence time intervals are assumed to be mutually independent and characterized by some arbitrary…

Condensed Matter · Physics 2009-11-10 Igor Goychuk , Peter Hanggi

This paper develops and analyzes an optimal-order semi-discrete scheme and its fully discrete finite element approximation for nonlinear stochastic elastic wave equations with multiplicative noise. A non-standard time-stepping scheme is…

Numerical Analysis · Mathematics 2025-04-08 Xiaobing Feng , Yukun Li , Liet Vo

In this paper, we present an online reinforcement learning algorithm for constrained Markov decision processes with a safety constraint. Despite the necessary attention of the scientific community, considering stochastic stopping time, the…

Machine Learning · Computer Science 2024-03-26 Abhijit Mazumdar , Rafal Wisniewski , Manuela L. Bujorianu

Scaled type Markov renewal processes generalize classical renewal processes: renewal times come from a one parameter family of probability laws and the sequence of the parameters is the trajectory of an ergodic Markov chain. Our primary…

Probability · Mathematics 2015-03-17 Zsolt Pajor-Gyulai , Domokos Szász

In this paper, we consider a class of stochastic optimal control problems with risk constraints that are expressed as bounded probabilities of failure for particular initial states. We present here a martingale approach that diffuses a risk…

Systems and Control · Computer Science 2015-07-09 Vu Anh Huynh , Leonid Kogan , Emilio Frazzoli

We study the statistical inference of nonlinear stochastic approximation algorithms utilizing a single trajectory of Markovian data. Our methodology has practical applications in various scenarios, such as Stochastic Gradient Descent (SGD)…

Statistics Theory · Mathematics 2023-02-21 Xiang Li , Jiadong Liang , Zhihua Zhang

We analyze the finite sample regret of a decreasing step size stochastic gradient algorithm. We assume correlated noise and use a perturbed Lyapunov function as a systematic approach for the analysis. Finally we analyze the escape time of…

Machine Learning · Computer Science 2024-10-14 George Yin , Vikram Krishnamurthy

We consider a finite state discrete time process X. Without loss of generality the finite state space can be identified with the set of unit vectors {e1, e2, . . . , eN} with ei = (0, . . . , 0, 1, 0, . . . , 0)0 2 RN. For a Markov chain…

Probability · Mathematics 2019-05-02 Robert J. Elliott

This paper addresses the problem of learning optimal policies for satisfying signal temporal logic (STL) specifications by agents with unknown stochastic dynamics. The system is modeled as a Markov decision process, in which the states…

Systems and Control · Computer Science 2016-09-26 Derya Aksaray , Austin Jones , Zhaodan Kong , Mac Schwager , Calin Belta
‹ Prev 1 8 9 10 Next ›