English
Related papers

Related papers: Theoretical guarantees for stochastic gradient sam…

200 papers

In this paper, the Milstein method is used to approximate invariant measures of stochastic differential equations with commutative noise. The decay rate of the transition probability kernel generated by the Milstein method to the unique…

Numerical Analysis · Mathematics 2019-01-28 Lihui Weng , Wei Liu

Understanding the space of probability measures on a metric space equipped with a Wasserstein distance is one of the fundamental questions in mathematical analysis. The Wasserstein metric has received a lot of attention in the machine…

Machine Learning · Computer Science 2021-03-02 Arijit Sehanobish , Neal Ravindra , David van Dijk

Wasserstein gradient flow (WGF) is a common method to perform optimization over the space of probability measures. While WGF is guaranteed to converge to a first-order stationary point, for nonconvex functionals the converged solution does…

Optimization and Control · Mathematics 2025-09-23 Naoya Yamamoto , Juno Kim , Taiji Suzuki

Variational inference, such as the mean-field (MF) approximation, requires certain conjugacy structures for efficient computation. These can impose unnecessary restrictions on the viable prior distribution family and further constraints on…

Statistics Theory · Mathematics 2023-09-11 Rentian Yao , Yun Yang

Let $X:=(X_t)_{t\geq 0}$ be an ergodic Markov process on $\real^d$, and $p>0$. We derive upper bounds of the $p$-Wasserstein distance between the invariant measure and the empirical measures of the Markov process $X$. For this we assume,…

Probability · Mathematics 2025-12-30 René L. Schilling , Jian Wang , Bingyao Wu , Jie-Xiang Zhu

We provide upper bounds of the expected Wasserstein distance between a probability measure and its empirical version, generalizing recent results for finite dimensional Euclidean spaces and bounded functional spaces. Such a generalization…

Statistics Theory · Mathematics 2020-01-29 Jing Lei

In this manuscript, we consider the Langevin dynamics on $\mathbb{R}^d$ with an overdamped vector field and driven by multiplicative Brownian noise of small amplitude $\sqrt{\epsilon}$, $\epsilon>0$. Under suitable assumptions on the vector…

Probability · Mathematics 2023-05-05 Gerardo Barrera

This paper studies the approximation of invariant measures of McKean-Vlasov dynamics with non-degenerate additive noise. While prior findings necessitated a strong monotonicity condition on the McKean-Vlasov process, we expand these results…

Probability · Mathematics 2024-01-24 Wenjing Cao , Kai Du

Comparing metric measure spaces (i.e. a metric space endowed with aprobability distribution) is at the heart of many machine learning problems. The most popular distance between such metric measure spaces is theGromov-Wasserstein (GW)…

Optimization and Control · Mathematics 2023-01-18 Thibault Séjourné , François-Xavier Vialard , Gabriel Peyré

Many machine learning problems can be formulated as minimax problems such as Generative Adversarial Networks (GANs), AUC maximization and robust estimation, to mention but a few. A substantial amount of studies are devoted to studying the…

Machine Learning · Computer Science 2021-07-14 Yunwen Lei , Zhenhuan Yang , Tianbao Yang , Yiming Ying

This paper is motivated by the problem of quantitatively bounding the convergence of adaptive control methods for stochastic systems to a stationary distribution. Such bounds are useful for analyzing statistics of trajectories and…

Optimization and Control · Mathematics 2021-10-19 Tyler Lekang , Andrew Lamperski

In this work, we describe a generic approach to show convergence with high probability for both stochastic convex and non-convex optimization with sub-Gaussian noise. In previous works for convex optimization, either the convergence is only…

Optimization and Control · Mathematics 2023-03-01 Zijian Liu , Ta Duy Nguyen , Thien Hang Nguyen , Alina Ene , Huy Lê Nguyen

In this work we provide performance guarantees for hypocoercive non-reversible MCMC samplers $X_t$ with invariant measure $\mu_*$; our results apply in particular to the Langevin equation, Hamiltonian Monte-Carlo, and the bouncy particle…

Probability · Mathematics 2025-10-13 Jeremiah Birrell , Luc Rey-Bellet

We study the approximation of a (finite) continuous-time Markov chain by a Markov chain on a reduced state space, and we provide formal error bounds for the approximated transient distributions in the Wasserstein distance. These bounds…

Probability · Mathematics 2025-12-19 Fabian Michel

Gaussian variational approximation is a popular methodology to approximate posterior distributions in Bayesian inference especially in high dimensional and large data settings. To control the computational cost while being able to capture…

Machine Learning · Computer Science 2021-04-07 Bingxin Zhou , Junbin Gao , Minh-Ngoc Tran , Richard Gerlach

The classical (overdamped) Langevin dynamics provide a natural algorithm for sampling from its invariant measure, which uniquely minimizes an energy functional over the space of probability measures, and which concentrates around the…

Probability · Mathematics 2023-09-26 Giovanni Conforti , Daniel Lacker , Soumik Pal

In inverse problems, many conditional generative models approximate the posterior measure by minimizing a distance between the joint measure and its learned approximation. While this approach also controls the distance between the posterior…

Machine Learning · Computer Science 2025-08-28 Jannis Chemseddine , Paul Hagemann , Gabriele Steidl , Christian Wald

We present a way to use Stein's method in order to bound the Wasserstein distance of order $2$ between two measures $\nu$ and $\mu$ supported on $\mathbb{R}^d$ such that $\mu$ is the reversible measure of a diffusion process. In order to…

Probability · Mathematics 2018-06-25 Thomas Bonis

The random splitting Langevin Monte Carlo could mitigate the first order bias in Langevin Monte Carlo with little extra work compared other high order schemes. We develop in this work an analysis framework for the sampling error under…

Numerical Analysis · Mathematics 2025-10-10 Lei Li , Chen Wang , Mengchao Wang

In this paper, we propose a novel technique to implement stochastic gradient methods, which are beneficial for learning from large datasets, through accelerated stochastic dynamics. A stochastic gradient method is based on mini-batch…

Machine Learning · Statistics 2016-05-04 Masayuki Ohzeki