English
Related papers

Related papers: A convergence rate for the entropic JKO scheme

200 papers

Distributionally Robust Optimization (DRO) has enabled to prove the equivalence between robustness and regularization in classification and regression, thus providing an analytical reason why regularization generalizes well in statistical…

Optimization and Control · Mathematics 2020-07-15 Esther Derman , Shie Mannor

The notion of entropy-regularized optimal transport, also known as Sinkhorn divergence, has recently gained popularity in machine learning and statistics, as it makes feasible the use of smoothed optimal transportation distances for data…

Statistics Theory · Mathematics 2019-11-05 Jérémie Bigot , Elsa Cazelles , Nicolas Papadakis

Score-based generative modeling with probability flow ordinary differential equations (ODEs) has achieved remarkable success in a variety of applications. While various fast ODE-based samplers have been proposed in the literature and…

Machine Learning · Statistics 2025-08-12 Xuefeng Gao , Lingjiong Zhu

Wasserstein distributionally robust optimization (WDRO) strengthens statistical learning under model uncertainty by minimizing the local worst-case risk within a prescribed ambiguity set. Although WDRO has been extensively studied in…

Machine Learning · Statistics 2025-11-12 Changyu Liu , Yuling Jiao , Junhui Wang , Jian Huang

Wasserstein distance plays increasingly important roles in machine learning, stochastic programming and image processing. Major efforts have been under way to address its high computational complexity, some leading to approximate or…

Machine Learning · Statistics 2019-06-26 Yujia Xie , Xiangfeng Wang , Ruijia Wang , Hongyuan Zha

The Sinkhorn "distance", a variant of the Wasserstein distance with entropic regularization, is an increasingly popular tool in machine learning and statistical inference. However, the time and memory requirements of standard algorithms for…

Machine Learning · Statistics 2021-11-16 Jason Altschuler , Francis Bach , Alessandro Rudi , Jonathan Niles-Weed

The recent remarkable progress of deep reinforcement learning (DRL) stands on regularization of policy for stable and efficient learning. A popular method, named proximal policy optimization (PPO), has been introduced for this purpose. PPO…

Machine Learning · Computer Science 2023-07-04 Taisuke Kobayashi

We present a discretization-free scalable framework for solving a large class of mass-conserving partial differential equations (PDEs), including the time-dependent Fokker-Planck equation and the Wasserstein gradient flow. The main…

Machine Learning · Computer Science 2023-11-15 Lingxiao Li , Samuel Hurault , Justin Solomon

We present a method to efficiently compute Wasserstein gradient flows. Our approach is based on a generalization of the back-and-forth method (BFM) introduced by Jacobs and L\'eger to solve optimal transport problems. We evolve the gradient…

Numerical Analysis · Mathematics 2020-11-17 Matt Jacobs , Wonjun Lee , Flavien Léger

In this work (Part I), we study three time-discretization procedures of the Dynamical Low-Rank Approximation (DLRA) of high-dimensional stochastic differential equations (SDEs). Specifically, we consider the Dynamically Orthogonal (DO)…

Numerical Analysis · Mathematics 2026-01-30 Yoshihito Kazashi , Fabio Nobile , Fabio Zoccolan

Stochastic approximation is a foundation for many algorithms found in machine learning and optimization. It is in general slow to converge: the mean square error vanishes as $O(n^{-1})$. A deterministic counterpart known as quasi-stochastic…

Optimization and Control · Mathematics 2024-03-26 Caio Kalil Lauand , Sean Meyn

Fix an irrational number $\alpha$. Let $X_1,X_2,\cdots$ be independent, identically distributed, integer-valued random variables with characteristic function $\varphi$, and let $S_n=\sum_{i=1}^n X_i$ be the partial sums. Consider the random…

Probability · Mathematics 2024-11-26 Bingyao Wu , Jie-Xiang Zhu

In 2012, Pflug and Pichler proved, under regularity assumptions, that the value function in Multistage Stochastic Programming (MSP) is Lipschitz continuous w.r.t. the Nested Distance, which is a distance between scenario trees (or discrete…

Optimization and Control · Mathematics 2021-07-22 Zheng Qu , Benoît Tran

This paper establishes the quantitative stability of invariant measures $\mu_{\alpha}$ for $\mathbb{R}^d$-valued ergodic stochastic differential equations driven by rotationally invariant multiplicative $\alpha$-stable processes with…

Probability · Mathematics 2025-09-17 Xinghu Jin , Xiaolong Zhang

This is an expository paper on the theory of gradient flows, and in particular of those PDEs which can be interpreted as gradient flows for the Wasserstein metric on the space of probability measures (a distance induced by optimal…

Analysis of PDEs · Mathematics 2016-09-14 Filippo Santambrogio

The nested distance builds on the Wasserstein distance to quantify the difference of stochastic processes, including also the information modelled by filtrations. The Sinkhorn divergence is a relaxation of the Wasserstein distance, which…

Optimization and Control · Mathematics 2021-02-11 Alois Pichler , Michael Weinhardt

We determine the convergence speed of a numerical scheme for approximating one-dimensional continuous strong Markov processes. The scheme is based on the construction of coin tossing Markov chains whose laws can be embedded into the process…

Probability · Mathematics 2020-08-26 Stefan Ankirchner , Thomas Kruse , Mikhail Urusov

The $2$-Wasserstein distance is sensitive to minor geometric differences between distributions, making it a very powerful dissimilarity metric. However, due to this sensitivity, a small outlier mass can also cause a significant increase in…

Machine Learning · Computer Science 2024-06-04 Sharath Raghvendra , Pouyan Shirzadian , Kaiyi Zhang

The approximation of invariant measures for nonlinear ergodic stochastic differential equations (SDEs) is a central problem in scientific computing, with important applications in stochastic sampling, physics, and ecology. We first propose…

Numerical Analysis · Mathematics 2025-11-18 Shan Huang , Xiaoyue Li

Making sense of Wasserstein distances between discrete measures in high-dimensional settings remains a challenge. Recent work has advocated a two-step approach to improve robustness and facilitate the computation of optimal transport, using…

Machine Learning · Computer Science 2019-09-04 François-Pierre Paty , Marco Cuturi