English
Related papers

Related papers: Finite Sample and Large Deviations Analysis of Sto…

200 papers

We present a theoretical analysis of some popular adaptive Stochastic Gradient Descent (SGD) methods in the small learning rate regime. Using the stochastic modified equations framework introduced by Li et al., we derive effective…

Machine Learning · Statistics 2025-09-29 Luca Callisti , Marco Romito , Francesco Triggiano

Motivated by applications to distributed optimization over networks and large-scale data processing in machine learning, we analyze the deterministic incremental aggregated gradient method for minimizing a finite sum of smooth functions…

Optimization and Control · Mathematics 2018-01-16 Mert Gurbuzbalaban , Asuman Ozdaglar , Pablo Parrilo

This paper presents new sufficient conditions for convergence and asymptotic or exponential stability of a stochastic discrete-time system, under which the constructed Lyapunov function always decreases in expectation along the system's…

Systems and Control · Computer Science 2019-06-05 Yuzhen Qin , Ming Cao , Brian D. O. Anderson

The large deviations analysis of solutions to stochastic differential equations and related processes is often based on approximation. The construction and justification of the approximations can be onerous, especially in the case where the…

Probability · Mathematics 2008-08-28 Amarjit Budhiraja , Paul Dupuis , Vasileios Maroulas

We revisit the finite time analysis of policy gradient methods in the one of the simplest settings: finite state and action MDPs with a policy class consisting of all stochastic policies and with exact gradient evaluations. There has been…

Machine Learning · Computer Science 2021-12-14 Jalaj Bhandari , Daniel Russo

Stochastic approximation is a powerful class of algorithms with celebrated success. However, a large body of previous analysis focuses on stochastic approximations driven by contractive operators, which is not applicable in some important…

Machine Learning · Computer Science 2025-11-21 Ethan Blaser , Shangtong Zhang

A scheme for stabilizing stochastic approximation iterates by adaptively scaling the step sizes is proposed and analyzed. This scheme leads to the same limiting differential equation as the original scheme and therefore has the same…

Probability · Mathematics 2010-07-28 Sameer Kamal

The graduated optimization approach is a method for finding global optimal solutions for nonconvex functions by using a function smoothing operation with stochastic noise. This paper makes three contributions regarding graduated…

Machine Learning · Computer Science 2026-01-27 Naoki Sato , Hideaki Iiduka

Many problems require to optimize empirical risk functions over large data sets. Gradient descent methods that calculate the full gradient in every descent step do not scale to such datasets. Various flavours of Stochastic Gradient Descent…

Machine Learning · Computer Science 2020-11-26 Francesco Cosentino , Harald Oberhauser , Alessandro Abate

In this paper, we consider asymptotic behaviors of multiscale multivalued stochastic systems with small noises. First of all, for general, fully coupled systems for multivalued stochastic differential equations of slow and fast motions with…

Probability · Mathematics 2025-09-30 Huijie Qiao

We consider a system of stochastic interacting particles in $\mathbb{R}^d$ and we describe large deviations asymptotics in a joint mean-field and small-noise limit. Precisely, a large deviations principle (LDP) is established for the…

Probability · Mathematics 2020-11-17 Carlo Orrieri

We study gradient descent (GD) with a constant stepsize for $\ell_2$-regularized logistic regression with linearly separable data. Classical theory suggests small stepsizes to ensure monotonic reduction of the optimization objective,…

Machine Learning · Statistics 2025-11-04 Jingfeng Wu , Pierre Marion , Peter Bartlett

Stochastic gradient methods are the workhorse (algorithms) of large-scale optimization problems in machine learning, signal processing, and other computational sciences and engineering. This paper studies Markov chain gradient descent, a…

Optimization and Control · Mathematics 2018-09-13 Tao Sun , Yuejiao Sun , Wotao Yin

We present a convergence analysis of a finite difference scheme for the time dependent partial different equation called gradient flow associated with the Rudin-Osher-Fatemi model. We devise an iterative algorithm to compute the solution of…

Numerical Analysis · Mathematics 2013-02-22 Qianying Hong , Ming-Jun Lai , Jingyue Wang

We study the problem of system identification for stochastic continuous-time dynamics, based on a single finite-length state trajectory. We present a method for estimating the possibly unstable open-loop matrix by employing properly…

Machine Learning · Statistics 2025-09-30 Reza Sadeghi Hafshejani , Mohamad Kazem Shirani Fradonbeh

A theoretical, and potentially also practical, problem with stochastic gradient descent is that trajectories may escape to infinity. In this note, we investigate uniform boundedness properties of iterates and function values along the…

Machine Learning · Computer Science 2022-06-23 Xiaoyu Wang , Mikael Johansson

In the machine learning literature stochastic gradient descent has recently been widely discussed for its purported implicit regularization properties. Much of the theory, that attempts to clarify the role of noise in stochastic gradient…

Machine Learning · Computer Science 2022-10-21 Alberto Lanconelli , Christopher S. A. Lauria

In this paper, we investigate the influence of noise giving an estimate of the gradient having a acute angle with the original. Noise amplitude has a relative model. The work offers both theoretical calculations and theorems, as well as…

Optimization and Control · Mathematics 2024-07-02 Artem Vasin

This paper considers an ergodic version of the bounded velocity follower problem, assuming that the decision maker lacks knowledge of the underlying system parameters and must learn them while simultaneously controlling. We propose…

Machine Learning · Statistics 2024-10-07 Stefan Ankirchner , Sören Christensen , Jan Kallsen , Philip Le Borne , Stefan Perko

Classical stochastic gradient methods for optimization rely on noisy gradient approximations that become progressively less accurate as iterates approach a solution. The large noise and small signal in the resulting gradients makes it…

Machine Learning · Computer Science 2017-04-10 Soham De , Abhay Yadav , David Jacobs , Tom Goldstein
‹ Prev 1 4 5 6 7 8 10 Next ›