English
Related papers

Related papers: Finite Time Analysis of Linear Two-timescale Stoch…

200 papers

The aim of this paper is to study the dynamical behavior of non-autonomous stochastic lattice systems with Markovian switching. We first show existence of an evolution system of measures of the stochastic system. We then study the pullback…

Dynamical Systems · Mathematics 2022-04-14 Dingshi Li , Yusen Lin , Zhe Pu

One of the most basic problems in reinforcement learning (RL) is policy evaluation: estimating the long-term return, i.e., value function, corresponding to a given fixed policy. The celebrated Temporal Difference (TD) learning algorithm…

Machine Learning · Computer Science 2025-02-10 Sreejeet Maity , Aritra Mitra

The theory of stochastic approximations form the theoretical foundation for studying convergence properties of many popular recursive learning algorithms in statistics, machine learning and statistical physics. Large deviations for…

Probability · Mathematics 2025-02-05 Henrik Hult , Adam Lindhe , Pierre Nyquist , Guo-Jhen Wu

We present a probabilistic model for stochastic iterative algorithms with the use case of optimization algorithms in mind. Based on this model, we present PAC-Bayesian generalization bounds for functions that are defined on the trajectory…

Machine Learning · Computer Science 2024-08-22 Michael Sucker , Peter Ochs

We study the almost sure convergence of the Stochastic Approximation algorithm to the fixed point $x^\star$ of a nonlinear operator under a negative drift condition and a general noise sequence with finite $p$-th moment for some $p > 1$.…

Optimization and Control · Mathematics 2026-02-23 Quang Dinh Thien Nguyen , Duc Anh Nguyen , Hoang Huy Nguyen , Siva Theja Maguluri

This paper studies fixed step-size stochastic approximation (SA) schemes, including stochastic gradient schemes, in a Riemannian framework. It is motivated by several applications, where geodesics can be computed explicitly, and their use…

Machine Learning · Statistics 2021-02-22 Alain Durmus , Pablo Jiménez , Éric Moulines , Salem Said

The phenomenological linear response theory of non-Markovian Stochastic Resonance (SR) is put forward for stationary two-state renewal processes. In terms of a derivation of a non-Markov regression theorem we evaluate the characteristic…

Statistical Mechanics · Physics 2007-05-23 Igor Goychuk , Peter Hanggi

We derive uniform all-time concentration bound of the type 'for all $n \geq n_0$ for some $n_0$' for TD(0) with linear function approximation. We work with online TD learning with samples from a single sample path of the underlying Markov…

Machine Learning · Computer Science 2026-01-13 Siddharth Chandak , Vivek S. Borkar

In this paper, we provide a unified analysis of temporal difference learning algorithms with linear function approximators by exploiting their connections to Markov jump linear systems (MJLS). We tailor the MJLS theory developed in the…

Machine Learning · Computer Science 2019-11-06 Bin Hu , Usman Ahmed Syed

Recent studies have shown that adaptive networks driven by simple local rules can organize into "critical" global steady states, providing another framework for self-organized criticality (SOC). We focus on the important convergence to…

Adaptation and Self-Organizing Systems · Physics 2013-05-30 Christian Kuehn

This work focuses on time-inhomogeneous Markov chains with two time scales. Our motivations stem from applications in reliability and dependability, queueing networks, financial engineering and manufacturing systems, where two-time-scale…

Probability · Mathematics 2007-05-23 George Yin , Hanqin Zhang

This paper compiles several aspects of the dynamics of stochastic approximation algorithms with Markov iterate-dependent noise when the iterates are not known to be stable beforehand. We achieve the same by extending the lock-in probability…

Dynamical Systems · Mathematics 2019-02-22 Prasenjit Karmakar , Shalabh Bhatnagar

We consider stochastic optimization problems with heavy-tailed noise with structured density. For such problems, we show that it is possible to get faster rates of convergence than $\mathcal{O}(K^{-2(\alpha - 1)/\alpha})$, when the…

Optimization and Control · Mathematics 2024-04-18 Nikita Puchkin , Eduard Gorbunov , Nikolay Kutuzov , Alexander Gasnikov

We study the problem of solving fixed-point equations for seminorm-contractive operators and establish foundational results on the non-asymptotic behavior of iterative algorithms in both deterministic and stochastic settings. Specifically,…

Machine Learning · Computer Science 2025-02-21 Zaiwei Chen , Sheng Zhang , Zhe Zhang , Shaan Ul Haque , Siva Theja Maguluri

We propose an algorithm to impute and forecast a time series by transforming the observed time series into a matrix, utilizing matrix estimation to recover missing values and de-noise observed entries, and performing linear regression to…

Machine Learning · Computer Science 2019-04-29 Anish Agarwal , Muhammad Jehangir Amjad , Devavrat Shah , Dennis Shen

In this paper, we present a methodology to estimate the parameters of stochastically contaminated models under two contamination regimes. In both regimes, we assume that the original process is a variable length Markov chain that is…

Methodology · Statistics 2017-02-23 Denise Duarte , Sokol Ndreca , Wecsley O. Prates

We consider stochastic optimization problems where data is drawn from a Markov chain. Existing methods for this setting crucially rely on knowing the mixing time of the chain, which in real-world applications is usually unknown. We propose…

Machine Learning · Computer Science 2023-07-14 Ron Dorfman , Kfir Y. Levy

We study the approximation of a Markov chain on a reduced state space, for both discrete- and continuous-time Markov chains. In this context, we extend the existing theory of formal error bounds for the approximated transient distributions.…

Probability · Mathematics 2025-02-12 Fabian Michel , Markus Siegle

We derive upper bounds for random design linear regression with dependent ($\beta$-mixing) data absent any realizability assumptions. In contrast to the strictly realizable martingale noise regime, no sharp instance-optimal non-asymptotics…

Machine Learning · Computer Science 2023-10-30 Ingvar Ziemann , Stephen Tu , George J. Pappas , Nikolai Matni

This work presents the first finite-time analysis for the last-iterate convergence of average-reward $Q$-learning with an asynchronous implementation. A key feature of the algorithm we study is the use of adaptive stepsizes, which serve as…

Machine Learning · Computer Science 2026-04-07 Zaiwei Chen , Phalguni Nanda