English
Related papers

Related papers: Minimizing $f$-Divergences by Interpolating Veloci…

200 papers

Training time-series forecasting models requires aligning the conditional distribution of model forecasts with that of the label sequence. The standard direct forecast (DF) approach resorts to minimizing the conditional negative…

Machine Learning · Computer Science 2026-04-14 Hao Wang , Licheng Pan , Yuan Lu , Zhixuan Chu , Xiaoxi Li , Shuting He , Zhichao Chen , Haoxuan Li , Qingsong Wen , Zhouchen Lin

In this paper, we establish sharp upper and lower bounds on the convergence rate of the empirical measures of point processes under the Wasserstein distance. To this end, we first introduce a new metric on the space of counting measures…

Statistics Theory · Mathematics 2026-04-28 Dongzhou Huang , Tianyi Jiang , Haonan Wang

Several issues in machine learning and inverse problems require to generate discrete data, as if sampled from a model probability distribution. A common way to do so relies on the construction of a uniform probability distribution over a…

Optimization and Control · Mathematics 2021-06-16 Quentin Merigot , Filippo Santambrogio , Clément Sarrazin

We prove that the sequence of marginals obtained from the iterations of the Sinkhorn algorithm or the iterative proportional fitting procedure (IPFP) on joint densities, converges to an absolutely continuous curve on the $2$-Wasserstein…

Probability · Mathematics 2026-04-21 Nabarun Deb , Young-Heon Kim , Soumik Pal , Geoffrey Schiebinger

We present a framework enabling variational data assimilation for gradient flows in general metric spaces, based on the minimizing movement (or Jordan-Kinderlehrer-Otto) approximation scheme. After discussing stability properties in the…

Numerical Analysis · Mathematics 2023-01-18 Jan-F. Pietschmann , Matthias Schlottbom

This paper presents a new gradient flow dissipation geometry over non-negative and probability measures. This is motivated by a principled construction that combines the unbalanced optimal transport and interaction forces modeled by…

Machine Learning · Computer Science 2024-11-01 Egor Gladin , Pavel Dvurechensky , Alexander Mielke , Jia-Jie Zhu

The Poisson-Nernst-Planck system of equations used to model ionic transport is interpreted as a gradient flow for the Wasserstein distance and a free energy in the space of probability measures with finite second moment. A variational…

Analysis of PDEs · Mathematics 2015-09-08 David Kinderlehrer , Léonard Monsaingeon , Xiang Xu

We propose conditional flows of the maximum mean discrepancy (MMD) with the negative distance kernel for posterior sampling and conditional generative modeling. This MMD, which is also known as energy distance, has several advantageous…

Wasserstein distributionally robust optimization offers a framework for model fitting in machine learning under potential shifts in the data distribution. We study a regularized variant of this problem in which entropic smoothing produces a…

Optimization and Control · Mathematics 2026-05-28 Tam Le

It is well-known that many diffusion equations can be recast as Wasserstein gradient flows. Moreover, in recent years, by modifying the Wasserstein distance appropriately, this technique has been transferred to further evolution equations…

Probability · Mathematics 2020-10-15 Kaveh Bashiri , Anton Bovier

Statistical inference can be performed by minimizing, over the parameter space, the Wasserstein distance between model distributions and the empirical distribution of the data. We study asymptotic properties of such minimum Wasserstein…

Methodology · Statistics 2019-05-13 Espen Bernton , Pierre E. Jacob , Mathieu Gerber , Christian P. Robert

This paper studies minimax optimization problems defined over infinite-dimensional function classes of overparameterized two-layer neural networks. In particular, we consider the minimax optimization problem stemming from estimating linear…

Machine Learning · Computer Science 2024-10-25 Yuchen Zhu , Yufeng Zhang , Zhaoran Wang , Zhuoran Yang , Xiaohong Chen

This paper introduces Wasserstein variational inference, a new form of approximate Bayesian inference based on optimal transport theory. Wasserstein variational inference uses a new family of divergences that includes both f-divergences and…

A simple model to handle the flow of people in emergency evacuation situations is considered: at every point x, the velocity U(x) that individuals at x would like to realize is given. Yet, the incompressibility constraint prevents this…

Analysis of PDEs · Mathematics 2010-02-04 Bertrand Maury , Aude Roudneff-Chupin , Filippo Santambrogio

We study policy gradient methods for continuous-action, entropy-regularized reinforcement learning through the lens of Wasserstein geometry. Starting from a Wasserstein proximal update, we derive Wasserstein Proximal Policy Gradient (WPPG)…

Machine Learning · Computer Science 2026-03-04 Zhaoyu Zhu , Shuhan Zhang , Rui Gao , Shuang Li

We propose a deterministic sampling framework using Score-Based Transport Modeling for sampling an unnormalized target density $\pi$ given only its score $\nabla \log \pi$. Our method approximates the Wasserstein gradient flow on…

Machine Learning · Computer Science 2025-10-21 Vasily Ilin , Peter Sushko , Jingwei Hu

Sampling a probability distribution with an unknown normalization constant is a fundamental problem in computational science and engineering. This task may be cast as an optimization problem over all probability measures, and an initial…

Machine Learning · Statistics 2024-09-12 Yifan Chen , Daniel Zhengyu Huang , Jiaoyang Huang , Sebastian Reich , Andrew M. Stuart

A nonlinear diffusion equation, interpreted as a Wasserstein gradient flow, is numerically solved in one space dimension using a higher-order minimizing movement scheme based on the BDF (backward differentiation formula) discretization. In…

Numerical Analysis · Mathematics 2015-09-02 Bertram Düring , Philipp Fuchs , Ansgar Jüngel

We study the efficacy and efficiency of deep generative networks for approximating probability distributions. We prove that neural networks can transform a low-dimensional source distribution to a distribution that is arbitrarily close to a…

Machine Learning · Computer Science 2023-12-05 Yunfei Yang , Zhen Li , Yang Wang

This paper presents a Wasserstein attraction approach for solving dynamic mass transport problems over networks. In the transport problem over networks, we start with a distribution over the set of nodes that needs to be "transported" to a…

Optimization and Control · Mathematics 2022-04-28 Ferran Arqué , César A. Uribe , Carlos Ocampo-Martinez