English
Related papers

Related papers: Theoretical guarantees for stochastic gradient sam…

200 papers

Over the last 25 years, techniques based on drift and minorization (d&m) have been mainstays in the convergence analysis of MCMC algorithms. However, results presented herein suggest that d&m may be less useful in the emerging area of…

Statistics Theory · Mathematics 2020-10-15 Qian Qin , James P. Hobert

We investigate contraction of the Wasserstein distances on $\mathbb{R}^d$ under Gaussian smoothing. It is well known that the heat semigroup is exponentially contractive with respect to the Wasserstein distances on manifolds of positive…

Probability · Mathematics 2020-12-15 Hong-Bin Chen , Jonathan Niles-Weed

The Wasserstein distance is a metric on a space of probability measures that has seen a surge of applications in statistics, machine learning, and applied mathematics. However, statistical aspects of Wasserstein distances are bottlenecked…

Probability · Mathematics 2022-03-02 Ziv Goldfeld , Kengo Kato , Sloan Nietert , Gabriel Rioux

In this paper we introduce a Wasserstein-type distance on the set of Gaussian mixture models. This distance is defined by restricting the set of possible coupling measures in the optimal transport problem to Gaussian mixture models. We…

Optimization and Control · Mathematics 2020-06-15 Julie Delon , Agnes Desolneux

Gaussian process (GP) regression is widely used for uncertainty quantification, yet the standard formulation assumes noise-free covariates. When inputs are measured with error, this errors-in-variables (EIV) setting can lead to…

Methodology · Statistics 2026-03-19 Hengrui Luo , Xiaoye S. Li , Yang Liu , Marcus Noack , Ji Qiang , Mark D. Risser

This paper proposes a stochastic gradient descent method with an adaptive Gaussian noise term for the global minimization of nearly convex functions, which are nonconvex and possess multiple strict local minimizers. The noise term,…

Optimization and Control · Mathematics 2025-08-05 Chenglong Bao , Liang Chen , Weizhi Shao

We present a novel approach to approximate Gaussian and mixture-of-Gaussians filtering. Our method relies on a variational approximation via a gradient-flow representation. The gradient flow is derived from a Kullback--Leibler discrepancy…

Computation · Statistics 2023-06-21 Adrien Corenflos , Hany Abdulsamad

We show how the infinitesimal exchangeable pairs approach to Stein's method combines naturally with the theory of Markov semigroups. We present a multivariate normal approximation theorem for functions of a random variable invariant with…

Probability · Mathematics 2025-10-01 David Grzybowski , Mark Meckes

Group-invariant probability distributions appear in many data-generative models in machine learning, such as graphs, point clouds, and images. In practice, one often needs to estimate divergences between such distributions. In this work, we…

Machine Learning · Computer Science 2026-02-05 Behrooz Tahmasebi , Stefanie Jegelka

Discrete time analogues of ergodic stochastic differential equations (SDEs) are one of the most popular and flexible tools for sampling high-dimensional probability measures. Non-asymptotic analysis in the $L^2$ Wasserstein distance of…

Probability · Mathematics 2019-10-11 Mateusz B. Majka , Aleksandar Mijatović , Lukasz Szpruch

Using quasi-Newton methods in stochastic optimization is not a trivial task given the difficulty of extracting curvature information from the noisy gradients. Moreover, pre-conditioning noisy gradient observations tend to amplify the noise.…

Optimization and Control · Mathematics 2024-04-02 Andre Carlon , Luis Espath , Raul Tempone

We study optimization problems whereby the optimization variable is a probability measure. Since the probability space is not a vector space, many classical and powerful methods for optimization (e.g., gradients) are of little help. Thus,…

Optimization and Control · Mathematics 2024-06-18 Nicolas Lanzetti , Antonio Terpin , Florian Dörfler

Generalization error bounds are essential to understanding machine learning algorithms. This paper presents novel expected generalization error upper bounds based on the average joint distribution between the output hypothesis and each…

Information Theory · Computer Science 2022-02-25 Gholamali Aminian , Yuheng Bu , Gregory Wornell , Miguel Rodrigues

We prove quantitative convergence rates at which discrete Langevin-like processes converge to the invariant distribution of a related stochastic differential equation. We study the setup where the additive noise can be non-Gaussian and…

Machine Learning · Computer Science 2020-11-20 Xiang Cheng , Dong Yin , Peter L. Bartlett , Michael I. Jordan

Motivated by approximation Bayesian computation using mean-field variational approximation and the computation of equilibrium in multi-species systems with cross-interaction, this paper investigates the composite geodesically convex…

Optimization and Control · Mathematics 2024-09-18 Rentian Yao , Xiaohui Chen , Yun Yang

We consider the problem of scalable sampling algorithms to fit Bayesian generalized linear mixed models on large datasets. Stochastic gradient Langevin dynamics, coupled with smooth re-parameterizations of variance parameters, produces…

Methodology · Statistics 2026-04-30 Youngsoo Baek , Samuel I. Berchuck

The computational complexity of MCMC methods for the exploration of complex probability measures is a challenging and important problem. A challenge of particular importance arises in Bayesian inverse problems where the target distribution…

Statistics Theory · Mathematics 2014-10-23 Sebastian J. Vollmer

This paper is devoted to the stochastic approximation of entropically regularized Wasserstein distances between two probability measures, also known as Sinkhorn divergences. The semi-dual formulation of such regularized optimal…

Statistics Theory · Mathematics 2024-12-10 Bernard Bercu , Jérémie Bigot

We quantify, uniformly over time and with high probability, the discrepancy between the predictions of a two-layer neural network trained by stochastic gradient descent (SGD) and their mean-field limit, for quadratic loss and ridge…

Neural and Evolutionary Computing · Computer Science 2026-03-03 Arnaud Guillin , Boris Nectoux , Paul Stos

This paper establishes the quantitative stability of invariant measures $\mu_{\alpha}$ for $\mathbb{R}^d$-valued ergodic stochastic differential equations driven by rotationally invariant multiplicative $\alpha$-stable processes with…

Probability · Mathematics 2025-09-17 Xinghu Jin , Xiaolong Zhang