English
Related papers

Related papers: Implicit Bias of the JKO Scheme

200 papers

Wasserstein distributionally robust optimization (\textsf{WDRO}) is a popular model to enhance the robustness of machine learning with ambiguous data. However, the complexity of \textsf{WDRO} can be prohibitive in practice since solving its…

Machine Learning · Computer Science 2023-05-10 Ruomin Huang , Jiawei Huang , Wenjie Liu , Hu Ding

Data-driven distributionally robust optimization is a recently emerging paradigm aimed at finding a solution that is driven by sample data but is protected against sampling errors. An increasingly popular approach, known as Wasserstein…

Optimization and Control · Mathematics 2022-07-20 Jonathan Yu-Meng Li , Tiantian Mao

In this note, we examine the forward-Euler discretization for simulating Wasserstein gradient flows. We provide two counter-examples showcasing the failure of this discretization even for a simple case where the energy functional is defined…

Machine Learning · Statistics 2024-06-13 Yewei Xu , Qin Li

We formulate well-posed continuous-time generative flows for learning distributions that are supported on low-dimensional manifolds through Wasserstein proximal regularizations of $f$-divergences. Wasserstein-1 proximal operators regularize…

Machine Learning · Statistics 2024-07-17 Hyemin Gu , Markos A. Katsoulakis , Luc Rey-Bellet , Benjamin J. Zhang

Recently, a Wasserstein analogue of the Cramer--Rao inequality has been developed using the Wasserstein information matrix (Otto metric). This inequality provides a lower bound on the Wasserstein variance of an estimator, which quantifies…

Statistics Theory · Mathematics 2025-08-26 Hayato Nishimori , Takeru Matsuda

Variational inference (VI) can be cast as an optimization problem in which the variational parameters are tuned to closely align a variational distribution with the true posterior. The optimization task can be approached through vanilla…

Machine Learning · Computer Science 2025-04-24 Dai Hai Nguyen , Tetsuya Sakurai , Hiroshi Mamitsuka

What is the optimal way to approximate a high-dimensional diffusion process by one in which the coordinates are independent? This paper presents a construction, called the \emph{independent projection}, which is optimal for two natural…

Probability · Mathematics 2024-10-10 Daniel Lacker

We consider the problem of sampling from a probability distribution $\pi$. It is well known that this can be written as an optimisation problem over the space of probability distribution in which we aim to minimise the Kullback--Leibler…

Methodology · Statistics 2026-02-11 Francesca R. Crucinio , Sahani Pathiraja

We present a Distributionally Robust Optimization (DRO) approach to estimate a robustified regression plane in a linear regression setting, when the observed samples are potentially contaminated with adversarially corrupted outliers. Our…

Machine Learning · Statistics 2018-05-14 Ruidi Chen , Ioannis Ch. Paschalidis

In recent years, Wasserstein Distributionally Robust Optimization (DRO) has garnered substantial interest for its efficacy in data-driven decision-making under distributional uncertainty. However, limited research has explored the…

Machine Learning · Computer Science 2025-10-01 Ahmad-Reza Ehyaei , Golnoosh Farnadi , Samira Samadi

We introduce the Wasserstein Transform (WT), a general unsupervised framework for updating distance structures on given data sets with the purpose of enhancing features and denoising. Our framework represents each data point by a…

Machine Learning · Computer Science 2026-04-14 Kun Jin , Facundo Mémoli , Zane Smith , Zhengchao Wan

Motivated by the 2D class averaging problem in single-particle cryo-electron microscopy (cryo-EM), we present a k-means algorithm based on a rotationally-invariant Wasserstein metric for images. Unlike existing methods that are based on…

Computer Vision and Pattern Recognition · Computer Science 2021-05-31 Rohan Rao , Amit Moscovich , Amit Singer

We prove the convergence of a Wasserstein gradient flow of a free energy in inhomogeneous media. Both the energy and media can depend on the spatial variable in a fast oscillatory manner. In particular, we show that the gradient-flow…

Analysis of PDEs · Mathematics 2025-08-19 Yuan Gao , Nung Kwan Yip

We show that several machine learning estimators, including square-root LASSO (Least Absolute Shrinkage and Selection) and regularized logistic regression can be represented as solutions to distributionally robust optimization (DRO)…

Statistics Theory · Mathematics 2020-10-22 Jose Blanchet , Yang Kang , Karthyek Murthy

The Swift--Hohenberg equation is a widely studied fourth-order model, originally proposed to describe hydrodynamic fluctuations. It admits an energy-dissipation law and, under suitable assumptions, bounded solutions. Many…

Numerical Analysis · Mathematics 2026-02-02 Yuki Yonekura , Daiki Iwade , Shun Sato , Takayasu Matsuo

Gromov-Wasserstein (GW) is a powerful tool to compare probability measures whose supports are in different metric spaces. GW suffers however from a computational drawback since it requires to solve a complex non-convex quadratic program. We…

Machine Learning · Statistics 2020-06-18 Tam Le , Nhat Ho , Makoto Yamada

Sampling a target probability distribution with an unknown normalization constant is a fundamental challenge in computational science and engineering. Recent work shows that algorithms derived by considering gradient flows in the space of…

Machine Learning · Statistics 2024-03-12 Yifan Chen , Daniel Zhengyu Huang , Jiaoyang Huang , Sebastian Reich , Andrew M Stuart

The theory of Wasserstein gradient flows in the space of probability measures has made an enormous progress over the last twenty years. It constitutes a unified and powerful framework in the study of dissipative partial differential…

Analysis of PDEs · Mathematics 2022-01-17 Daniel Adams , Manh Hong Duong , Goncalo dos Reis

We prove the convergence of a modified Jordan--Kinderlehrer--Otto scheme to a solution to the Fokker--Planck equation in $\Omega \Subset \mathbb R^d$ with general -- strictly positive and temporally constant -- Dirichlet boundary…

Analysis of PDEs · Mathematics 2025-12-12 Filippo Quattrocchi

Many tasks in machine learning and signal processing can be solved by minimizing a convex function of a measure. This includes sparse spikes deconvolution or training a neural network with a single hidden layer. For these problems, we study…

Optimization and Control · Mathematics 2018-10-30 Lenaic Chizat , Francis Bach
‹ Prev 1 8 9 10 Next ›