English
Related papers

Related papers: Neural Wasserstein Gradient Flows for Maximum Mean…

200 papers

We construct a Wasserstein gradient flow of the maximum mean discrepancy (MMD) and study its convergence properties. The MMD is an integral probability metric defined for a reproducing kernel Hilbert space (RKHS), and serves as a metric on…

Machine Learning · Statistics 2019-12-04 Michael Arbel , Anna Korba , Adil Salim , Arthur Gretton

We give a comprehensive description of Wasserstein gradient flows of maximum mean discrepancy (MMD) functionals $\mathcal F_\nu := \text{MMD}_K^2(\cdot, \nu)$ towards given target measures $\nu$ on the real line, where we focus on the…

Analysis of PDEs · Mathematics 2025-12-05 Richard Duong , Viktor Stein , Robert Beinert , Johannes Hertrich , Gabriele Steidl

The aim of this paper is twofold. Based on the geometric Wasserstein tangent space, we first introduce Wasserstein steepest descent flows. These are locally absolutely continuous curves in the Wasserstein space whose tangent vectors point…

Optimization and Control · Mathematics 2024-02-06 Johannes Hertrich , Manuel Gräf , Robert Beinert , Gabriele Steidl

Maximum mean discrepancy (MMD) flows suffer from high computational costs in large scale computations. In this paper, we show that MMD flows with Riesz kernels $K(x,y) = - \|x-y\|^r$, $r \in (0,2)$ have exceptional properties which allow…

Machine Learning · Computer Science 2024-02-21 Johannes Hertrich , Christian Wald , Fabian Altekrüger , Paul Hagemann

We study the quantitative convergence of Wasserstein gradient flows of Kernel Mean Discrepancy (KMD) (also known as Maximum Mean Discrepancy (MMD)) functionals. Our setting covers in particular the training dynamics of shallow neural…

Analysis of PDEs · Mathematics 2026-03-03 Lénaïc Chizat , Maria Colombo , Roberto Colombo , Xavier Fernández-Real

We introduce a (de)-regularization of the Maximum Mean Discrepancy (DrMMD) and its Wasserstein gradient flow. Existing gradient flows that transport samples from source distribution to target distribution with only target samples, either…

We propose conditional flows of the maximum mean discrepancy (MMD) with the negative distance kernel for posterior sampling and conditional generative modeling. This MMD, which is also known as energy distance, has several advantageous…

This study focuses on a Wasserstein-type gradient flow, which represents an optimization process of a continuous model of a Deep Neural Network (DNN). First, we establish the existence of a minimizer for an average loss of the model under…

Machine Learning · Computer Science 2024-04-16 Noboru Isobe

This paper provides results on Wasserstein gradient flows between measures on the real line. Utilizing the isometric embedding of the Wasserstein space $\mathcal P_2(\mathbb R)$ into the Hilbert space $L_2((0,1))$, Wasserstein gradient…

Optimization and Control · Mathematics 2024-08-13 Johannes Hertrich , Robert Beinert , Manuel Gräf , Gabriele Steidl

We consider Wasserstein gradient flows of maximum mean discrepancy (MMD) functionals $\text{MMD}_K^2(\cdot, \nu)$ for positive and negative distance kernels $K(x,y) := \pm |x-y|$ and given target measures $\nu$ on $\mathbb{R}$. Since in one…

Analysis of PDEs · Mathematics 2025-05-06 Richard Duong , Nicolaj Rux , Viktor Stein , Gabriele Steidl

Negative distance kernels $K(x,y) := - \|x-y\|$ were used in the definition of maximum mean discrepancies (MMDs) in statistics and lead to favorable numerical results in various applications. In particular, so-called slicing techniques for…

Machine Learning · Statistics 2025-10-23 Nicolaj Rux , Michael Quellmalz , Gabriele Steidl

Commonly used $f$-divergences of measures, e.g., the Kullback-Leibler divergence, are subject to limitations regarding the support of the involved measures. A remedy is regularizing the $f$-divergence by a squared maximum mean discrepancy…

Machine Learning · Statistics 2025-04-14 Viktor Stein , Sebastian Neumayer , Nicolaj Rux , Gabriele Steidl

The purpose of this paper is to answer a few open questions in the interface of kernel methods and PDE gradient flows. Motivated by recent advances in machine learning, particularly in generative modeling and sampling, we present a rigorous…

Machine Learning · Statistics 2024-10-29 Jia-Jie Zhu , Alexander Mielke

In this paper, we propose a novel numerical scheme to optimize the gradient flows for learning energy-based models (EBMs). From a perspective of physical simulation, we redefine the problem of approximating the gradient flow utilizing…

Computer Vision and Pattern Recognition · Computer Science 2023-05-01 Yang Wu , Pengxu Wei , Liang Lin

We study the convergence of gradient flow for the training of deep neural networks. If Residual Neural Networks are a popular example of very deep architectures, their training constitutes a challenging optimization problem due notably to…

Machine Learning · Computer Science 2025-07-22 Raphaël Barboni , Gabriel Peyré , François-Xavier Vialard

We propose a gradient flow procedure for generative modeling by transporting particles from an initial source distribution to a target distribution, where the gradient field on the particles is given by a noise-adaptive Wasserstein Gradient…

Machine Learning · Computer Science 2024-05-14 Alexandre Galashov , Valentin de Bortoli , Arthur Gretton

We reveal a precise mathematical framework about a new family of generative models which we call Gradient Flow Drifting. With this framework, we prove an equivalence between the recently proposed Drifting Model and the Wasserstein gradient…

Machine Learning · Computer Science 2026-03-12 Jiarui Cao , Zixuan Wei , Yuxin Liu

We consider the maximum mean discrepancy ($\mathrm{MMD}$) GAN problem and propose a parametric kernelized gradient flow that mimics the min-max game in gradient regularized $\mathrm{MMD}$ GAN. We show that this flow provides a descent…

Machine Learning · Computer Science 2020-11-05 Youssef Mroueh , Truyen Nguyen

This paper studies minimax optimization problems defined over infinite-dimensional function classes of overparameterized two-layer neural networks. In particular, we consider the minimax optimization problem stemming from estimating linear…

Machine Learning · Computer Science 2024-10-25 Yuchen Zhu , Yufeng Zhang , Zhaoran Wang , Zhuoran Yang , Xiaohong Chen

We provide a numerical analysis and computation of neural network projected schemes for approximating one dimensional Wasserstein gradient flows. We approximate the Lagrangian mapping functions of gradient flows by the class of two-layer…

Numerical Analysis · Mathematics 2024-02-27 Xinzhe Zuo , Jiaxi Zhao , Shu Liu , Stanley Osher , Wuchen Li
‹ Prev 1 2 3 10 Next ›