English
Related papers

Related papers: A New Robust Partial $p$-Wasserstein-Based Metric …

200 papers

Wasserstein distributionally robust optimization (WDRO) strengthens statistical learning under model uncertainty by minimizing the local worst-case risk within a prescribed ambiguity set. Although WDRO has been extensively studied in…

Machine Learning · Statistics 2025-11-12 Changyu Liu , Yuling Jiao , Junhui Wang , Jian Huang

In this paper we introduce some recent progresses on the convergence rate in Wasserstein distance for empirical measures of Markov processes. For diffusion processes on compact manifolds possibly with reflecting or killing boundary…

Probability · Mathematics 2025-07-22 Feng-Yu Wang

We obtain explicit $p$-Wasserstein distance error bounds between the distribution of the multi-parameter MLE and the multivariate normal distribution. Our general bounds are given for possibly high-dimensional, independent and identically…

Statistics Theory · Mathematics 2021-12-28 Andreas Anastasiou , Robert E. Gaunt

Issued from Optimal Transport, the Wasserstein distance has gained importance in Machine Learning due to its appealing geometrical properties and the increasing availability of efficient approximations. In this work, we consider the problem…

Machine Learning · Statistics 2022-02-21 Guillaume Staerman , Pierre Laforgue , Pavlo Mozharovskyi , Florence d'Alché-Buc

We introduce a new version of the KL-divergence for Gaussian distributions which is based on Wasserstein geometry and referred to as WKL-divergence. We show that this version is consistent with the geometry of the sample space ${\Bbb R}^n$.…

Statistics Theory · Mathematics 2026-05-29 Adwait Datar , Nihat Ay

The Sliced Gromov-Wasserstein (SGW) distance, aiming to relieve the computational cost of solving a non-convex quadratic program that is the Gromov-Wasserstein distance, utilizes projecting directions sampled uniformly from unit…

Machine Learning · Statistics 2025-07-18 Dhruv Sarkar , Aprameyo Chakrabartty , Anish Chakrabarty , Swagatam Das

We introduce a distributionally robust maximum likelihood estimation model with a Wasserstein ambiguity set to infer the inverse covariance matrix of a $p$-dimensional Gaussian random vector from $n$ independent samples. The proposed model…

Optimization and Control · Mathematics 2018-05-21 Viet Anh Nguyen , Daniel Kuhn , Peyman Mohajerin Esfahani

This paper presents a distance-based discriminative framework for learning with probability distributions. Instead of using kernel mean embeddings or generalized radial basis kernels, we introduce embeddings based on dissimilarity of…

Machine Learning · Computer Science 2018-11-16 Alain Rakotomamonjy , Abraham Traoré , Maxime Berar , Rémi Flamary , Nicolas Courty

Understanding proper distance measures between distributions is at the core of several learning tasks such as generative models, domain adaptation, clustering, etc. In this work, we focus on mixture distributions that arise naturally in…

Machine Learning · Computer Science 2019-10-30 Yogesh Balaji , Rama Chellappa , Soheil Feizi

We consider multiperiod stochastic control problems with non-parametric uncertainty on the underlying probabilistic model. We derive a new metric on the space of probability measures, called the adapted $(p, \infty)$--Wasserstein distance…

Optimization and Control · Mathematics 2024-11-01 Ruslan Mirmominov , Johannes Wiesel

Since the introduction of the Sliced Wasserstein distance in the literature, its simplicity and efficiency have made it one of the most interesting surrogate for the Wasserstein distance in image processing and machine learning. However,…

Optimization and Control · Mathematics 2025-08-05 Eloi Tanguy , Laetitia Chapel , Julie Delon

To measure the similarity of documents, the Wasserstein distance is a powerful tool, but it requires a high computational cost. Recently, for fast computation of the Wasserstein distance, methods for approximating the Wasserstein distance…

Machine Learning · Computer Science 2021-07-26 Yuki Takezawa , Ryoma Sato , Makoto Yamada

For $\ell\colon \mathbb{R}^d \to [0,\infty)$ we consider the sequence of probability measures $\left(\mu_n\right)_{n \in \mathbb{N}}$, where $\mu_n$ is determined by a density that is proportional to $\exp(-n\ell)$. We allow for infinitely…

Probability · Mathematics 2023-12-11 Mareike Hasenpflug , Daniel Rudolf , Björn Sprungk

Computing the quadratic transportation metric (also called the $2$-Wasserstein distance or root mean square distance) between two point clouds, or, more generally, two discrete distributions, is a fundamental problem in machine learning,…

Data Structures and Algorithms · Computer Science 2018-12-18 Jason Altschuler , Francis Bach , Alessandro Rudi , Jonathan Weed

Nonparametric two sample or homogeneity testing is a decision theoretic problem that involves identifying differences between two random variables without making parametric assumptions about their underlying distributions. The literature is…

Statistics Theory · Mathematics 2015-10-14 Aaditya Ramdas , Nicolas Garcia , Marco Cuturi

In this paper, we investigate Dimensionality reduction (DR) maps in an information retrieval setting from a quantitative topology point of view. In particular, we show that no DR maps can achieve perfect precision and perfect recall…

Machine Learning · Statistics 2018-11-02 Kry Yik Chau Lui , Gavin Weiguang Ding , Ruitong Huang , Robert J. McCann

Markov decision processes (MDPs) are known to be sensitive to parameter specification. Distributionally robust MDPs alleviate this issue by allowing for \emph{ambiguity sets} which give a set of possible distributions over parameter sets.…

Optimization and Control · Mathematics 2021-05-05 Julien Grand-Clément , Christian Kroer

The Wasserstein metric is an important measure of distance between probability distributions, with applications in machine learning, statistics, probability theory, and data analysis. This paper provides upper and lower bounds on…

Statistics Theory · Mathematics 2019-11-11 Shashank Singh , Barnabás Póczos

Wasserstein Discriminant Analysis (WDA) is a new supervised method that can improve classification of high-dimensional data by computing a suitable linear map onto a lower dimensional subspace. Following the blueprint of classical Linear…

Machine Learning · Statistics 2018-09-21 Rémi Flamary , Marco Cuturi , Nicolas Courty , Alain Rakotomamonjy

We introduce a general framework for analyzing data modeled as parameterized families of networks. Building on a Gromov-Wasserstein variant of optimal transport, we define a family of parameterized Gromov-Wasserstein distances for comparing…

Machine Learning · Statistics 2025-09-29 Mario Gómez , Guanqun Ma , Tom Needham , Bei Wang