English
Related papers

Related papers: Geometric ergodicity of SGLD via reflection coupli…

200 papers

We study the implicit Langevin Monte Carlo (iLMC) method, which simulates the overdamped Langevin equation via an implicit iteration rule. In many applications, iLMC is favored over other explicit schemes such as the (explicit) Langevin…

Numerical Analysis · Mathematics 2025-11-07 Lei Li , Jian-Guo Liu , Yuliang Wang

The current interpretation of stochastic gradient descent (SGD) as a stochastic process lacks generality in that its numerical scheme restricts continuous-time dynamics as well as the loss function and the distribution of gradient noise. We…

Machine Learning · Statistics 2019-11-21 Soma Yokoi , Issei Sato

The so-called SAGA-LD algorithm is used for efficient sampling from high-dimensional distributions in machine learning. Its intricate dynamics resists standard approaches of Markov chain theory. We prove, using a model-specific method, that…

Probability · Mathematics 2026-04-15 Miklós Rásonyi

We consider the task of sampling with respect to a log concave probability distribution. The potential of the target distribution is assumed to be composite, \textit{i.e.}, written as the sum of a smooth convex term, and a nonsmooth convex…

Machine Learning · Statistics 2021-02-23 Adil Salim , Peter Richtárik

We provide a framework to analyze the convergence of discretized kinetic Langevin dynamics for $M$-$\nabla$Lipschitz, $m$-convex potentials. Our approach gives convergence rates of $\mathcal{O}(m/M)$, with explicit stepsize restrictions,…

Numerical Analysis · Mathematics 2024-05-24 Benedict Leimkuhler , Daniel Paulin , Peter A. Whalley

Boundary discontinuity and its inconsistency to the final detection metric have been the bottleneck for rotating detection regression loss design. In this paper, we propose a novel regression loss based on Gaussian Wasserstein distance as a…

Computer Vision and Pattern Recognition · Computer Science 2022-04-19 Xue Yang , Junchi Yan , Qi Ming , Wentao Wang , Xiaopeng Zhang , Qi Tian

Sampling with Markov chain Monte Carlo methods often amounts to discretizing some continuous-time dynamics with numerical integration. In this paper, we establish the convergence rate of sampling algorithms obtained by discretizing smooth…

Machine Learning · Statistics 2020-02-04 Xuechen Li , Denny Wu , Lester Mackey , Murat A. Erdogdu

Machine learning has made tremendous progress in recent years, with models matching or even surpassing humans on a series of specialized tasks. One key element behind the progress of machine learning in recent years has been the ability to…

Machine Learning · Computer Science 2020-06-30 Giorgi Nadiradze , Ilia Markov , Bapi Chatterjee , Vyacheslav Kungurtsev , Dan Alistarh

In this article, we discuss ergodicity properties of a diffusion process given through an It\^{o} stochastic differential equation. We identify conditions on the drift and diffusion coefficients which result in sub-geometric ergodicity of…

Probability · Mathematics 2020-06-03 Petra Lazić , Nikola Sandrić

Stochastic gradient descent (SGD) method is popular for solving non-convex optimization problems in machine learning. This work investigates SGD from a viewpoint of graduated optimization, which is a widely applied approach for non-convex…

Optimization and Control · Mathematics 2023-08-15 Da Li , Jingjing Wu , Qingrun Zhang

We propose an interacting contour stochastic gradient Langevin dynamics (ICSGLD) sampler, an embarrassingly parallel multiple-chain contour stochastic gradient Langevin dynamics (CSGLD) sampler with efficient interactions. We show that…

Machine Learning · Statistics 2022-02-22 Wei Deng , Siqi Liang , Botao Hao , Guang Lin , Faming Liang

We describe an implicit general relativistic hydrodynamics code. The evolution equations are formulated in comoving coordinates. A conservative finite differencing of the Einstein equations is outlined, and artificial viscosity and…

Astrophysics · Physics 2010-05-12 Matthias Liebendoerfer , Stephan Rosswog , Friedrich-Karl Thielemann

We study the contraction in Wasserstein distance of the coordinate ascent variational inference algorithm. This is shown to hold under a transport-information inequality at the fixed points and a functional smoothness condition. The results…

Machine Learning · Statistics 2026-05-29 Rocco Caprio , Adrien Corenflos , Sam Power

Training neural networks requires optimizing a loss function that may be highly irregular, and in particular neither convex nor smooth. Popular training algorithms are based on stochastic gradient descent with momentum (SGDM), for which…

Machine Learning · Computer Science 2026-03-17 Qinzi Zhang , Ashok Cutkosky

Distribution matching is central to many vision and graphics tasks, where the widely used Wasserstein distance is too costly to compute for high dimensional distributions. The Sliced Wasserstein Distance (SWD) offers a scalable alternative,…

Graphics · Computer Science 2025-10-02 Mark Boss , Andreas Engelhardt , Simon Donné , Varun Jampani

In any Markov chain Monte Carlo analysis, rapid convergence of the chain to its target probability distribution is of practical and theoretical importance. A chain that converges at a geometric rate is geometrically ergodic. In this paper,…

Computation · Statistics 2012-10-05 Alicia A. Johnson , Owen Burbank

LocalSGD and SCAFFOLD are widely used methods in distributed stochastic optimization, with numerous applications in machine learning, large-scale data processing, and federated learning. However, rigorously establishing their theoretical…

Optimization and Control · Mathematics 2025-02-25 Ruichen Luo , Sebastian U Stich , Samuel Horváth , Martin Takáč

Stochastic Gradient Descent Langevin Dynamics (SGLD) algorithms, which add noise to the classic gradient descent, are known to improve the training of neural networks in some cases where the neural network is very deep. In this paper we…

Computational Finance · Quantitative Finance 2023-01-16 Pierre Bras , Gilles Pagès

In this paper, we study the performance of a large family of SGD variants in the smooth nonconvex regime. To this end, we propose a generic and flexible assumption capable of accurate modeling of the second moment of the stochastic…

Optimization and Control · Mathematics 2020-06-15 Zhize Li , Peter Richtárik

We consider stochastic optimization problems where the objective depends on some parameter, as commonly found in hyperparameter optimization for instance. We investigate the behavior of the derivatives of the iterates of Stochastic Gradient…

Optimization and Control · Mathematics 2024-11-21 Franck Iutzeler , Edouard Pauwels , Samuel Vaiter
‹ Prev 1 8 9 10 Next ›