English
Related papers

Related papers: Finite-Time Error Bounds For Linear Stochastic App…

200 papers

Stochastic iterative algorithms, including stochastic gradient descent (SGD) and stochastic gradient Langevin dynamics (SGLD), are widely utilized for optimization and sampling in large-scale and high-dimensional problems in machine…

Machine Learning · Statistics 2025-01-22 Xiaoyu Wang , Mikolaj J. Kasprzak , Jeffrey Negrea , Solesne Bourguin , Jonathan H. Huggins

In this paper, we study the dynamics of temporal difference learning with neural network-based value function approximation over a general state space, namely, \emph{Neural TD learning}. We consider two practically used algorithms,…

Machine Learning · Computer Science 2021-08-09 Semih Cayci , Siddhartha Satpathi , Niao He , R. Srikant

We present a systematic study of moment evolution in multidimensional stochastic difference systems, focusing on characterizing systems whose low-order moments diverge in the neighborhood of a stable fixed point. We consider systems with a…

Mathematical Physics · Physics 2009-11-10 Dennis M. Wilkinson

We study a decentralized variant of stochastic approximation, a data-driven approach for finding the root of an operator under noisy measurements. A network of agents, each with its own operator and data observations, cooperatively find the…

Machine Learning · Computer Science 2022-06-17 Sihan Zeng , Thinh T. Doan , Justin Romberg

Stochastic approximation is a class of algorithms that update a vector iteratively, incrementally, and stochastically, including, e.g., stochastic gradient descent and temporal difference learning. One fundamental challenge in analyzing a…

Machine Learning · Computer Science 2025-11-06 Shuze Daniel Liu , Shuhang Chen , Shangtong Zhang

We establish new conditions for obtaining uniform bounds on the moments of discrete-time stochastic processes. Our results require a weak negative drift criterion along with a state-dependent restriction on the sizes of the one-step jumps…

Probability · Mathematics 2022-06-02 Arnab Ganguly , Debasish Chatterjee

ODE solvers with randomly sampled timestep sizes appear in the context of chaotic dynamical systems, differential equations with low regularity, and, implicitly, in stochastic optimisation. In this work, we propose and study the stochastic…

Numerical Analysis · Mathematics 2024-08-05 Jonas Latz

We study the approximation of a Markov chain on a reduced state space, for both discrete- and continuous-time Markov chains. In this context, we extend the existing theory of formal error bounds for the approximated transient distributions.…

Probability · Mathematics 2025-02-12 Fabian Michel , Markus Siegle

This paper presents new sufficient conditions for convergence and asymptotic or exponential stability of a stochastic discrete-time system, under which the constructed Lyapunov function always decreases in expectation along the system's…

Systems and Control · Computer Science 2019-06-05 Yuzhen Qin , Ming Cao , Brian D. O. Anderson

In this paper, we analyze the finite sample complexity of stochastic system identification using modern tools from machine learning and statistics. An unknown discrete-time linear system evolves over time under Gaussian noise without…

Machine Learning · Computer Science 2019-03-22 Anastasios Tsiamis , George J. Pappas

We recently proposed a method for estimation of states and parameters in stochastic differential equations, which included intermediate time points between observations and used the Laplace approximation to integrate out these intermediate…

Probability · Mathematics 2025-04-01 Uffe Høgsbro Thygesen

The theory of stochastic approximations form the theoretical foundation for studying convergence properties of many popular recursive learning algorithms in statistics, machine learning and statistical physics. Large deviations for…

Probability · Mathematics 2025-02-05 Henrik Hult , Adam Lindhe , Pierre Nyquist , Guo-Jhen Wu

We revisit the convergence analysis of constant stepsize stochastic approximation (SA) with decision-dependent Markovian noise, with a focus on characterizing the stationary bias against the root of the mean-field equation. We first…

Optimization and Control · Mathematics 2026-04-16 Hadi Hadavi , Wenlong Mou , Sergey Samsonov , Hoi-To Wai

In this paper, we present a Longstaff-Schwartz-type algorithm for optimal stopping time problems based on the Brownian motion filtration. The algorithm is based on Le\~ao, Ohashi and Russo and, in contrast to previous works, our methodology…

Computational Finance · Quantitative Finance 2019-12-05 Sérgio C. Bezerra , Alberto Ohashi , Francesco Russo , Francys de Souza

Two-timescale stochastic approximation (TTSA) is among the most general frameworks for iterative stochastic algorithms. This includes well-known stochastic optimization methods such as SGD variants and those designed for bilevel or minimax…

Machine Learning · Statistics 2024-02-15 Jie Hu , Vishwaraj Doshi , Do Young Eun

We investigate the finite-time convergence properties of Temporal Difference (TD) learning with linear function approximation, a cornerstone algorithm in the field of reinforcement learning. We are interested in the so-called ``robust''…

Machine Learning · Computer Science 2025-09-26 Wei-Cheng Lee , Francesco Orabona

Stochastic time-varying optimization is an integral part of learning in which the shape of the function changes over time in a non-deterministic manner. This paper considers multiple models of stochastic time variation and analyzes the…

Optimization and Control · Mathematics 2023-02-23 Ali Yekkehkhany , Han Feng , Donghao Ying , Javad Lavaei

This paper studies the stability properties of stochastic differential equations subject to persistent noise (including the case of additive noise), which is noise that is present even at the equilibria of the underlying differential…

Dynamical Systems · Mathematics 2015-01-22 D. Mateos-Núñez , J. Cortés

We study systems on time scales that are generalizations of classical differential or difference equations. In this paper we consider linear systems and their small nonlinear perturbations. In terms of time scales and of eigenvalues of…

Dynamical Systems · Mathematics 2016-06-07 Sergey Kryzhevich , Alexander Nazarov

We derive uniform all-time concentration bound of the type 'for all $n \geq n_0$ for some $n_0$' for TD(0) with linear function approximation. We work with online TD learning with samples from a single sample path of the underlying Markov…

Machine Learning · Computer Science 2026-01-13 Siddharth Chandak , Vivek S. Borkar
‹ Prev 1 3 4 5 6 7 10 Next ›