English
Related papers

Related papers: Mean and Variance Estimation Complexity in Arbitra…

200 papers

This paper is focused on the study of entropic regularization in optimal transport as a smoothing method for Wasserstein estimators, through the prism of the classical tradeoff between approximation and estimation errors in statistics.…

Machine Learning · Statistics 2024-10-30 Jérémie Bigot , Paul Freulon , Boris P. Hejblum , Arthur Leclaire

The mean field variational inference (MFVI) formulation restricts the general Bayesian inference problem to the subspace of product measures. We present a framework to analyze MFVI algorithms, which is inspired by a similar development for…

Machine Learning · Statistics 2022-10-21 Soumyadip Ghosh , Yingdong Lu , Tomasz Nowicki , Edith Zhang

We study the fixed-support Wasserstein barycenter problem (FS-WBP), which consists in computing the Wasserstein barycenter of $m$ discrete probability measures supported on a finite metric space of size $n$. We show first that the…

Computational Complexity · Computer Science 2022-06-07 Tianyi Lin , Nhat Ho , Xi Chen , Marco Cuturi , Michael I. Jordan

In classical statistics and distribution testing, it is often assumed that elements can be sampled from some distribution $P$, and that when an element $x$ is sampled, the probability $P$ of sampling $x$ is also known. Recent work in…

Data Structures and Algorithms · Computer Science 2022-08-03 Talya Eden , Jakob Bæk Tejs Houen , Shyam Narayanan , Will Rosenbaum , Jakub Tětek

Wasserstein barycenters provide a geometric notion of the weighted average of probability measures based on optimal transport. In this paper, we present a scalable algorithm to compute Wasserstein-2 barycenters given sample access to the…

Machine Learning · Computer Science 2022-01-02 Alexander Korotin , Lingxiao Li , Justin Solomon , Evgeny Burnaev

We revisit Markowitz's mean-variance portfolio selection model by considering a distributionally robust version, where the region of distributional uncertainty is around the empirical measure and the discrepancy between probability measures…

Methodology · Statistics 2018-02-15 Jose Blanchet , Lin Chen , Xun Yu Zhou

This paper studies the problem of computing a linear approximation of quadratic Wasserstein distance $W_2$. In particular, we compute an approximation of the negative homogeneous weighted Sobolev norm whose connection to Wasserstein…

Numerical Analysis · Mathematics 2022-03-02 Philip Greengard , Jeremy G. Hoskins , Nicholas F. Marshall , Amit Singer

Multi-marginal optimal transport enables one to compare multiple probability measures, which increasingly finds application in multi-task learning problems. One practical limitation of multi-marginal transport is computational scalability…

To consider model uncertainty in global Fr\'{e}chet regression and improve density response prediction, we propose a frequentist model averaging method. The weights are chosen by minimizing a cross-validation criterion based on Wasserstein…

Methodology · Statistics 2023-09-06 Xingyu Yan , Xinyu Zhang , Peng Zhao

This paper presents a computational framework for the concise encoding of an ensemble of persistence diagrams, in the form of weighted Wasserstein barycenters [100], [102] of a dictionary of atom diagrams. We introduce a multi-scale…

Machine Learning · Computer Science 2023-09-18 Keanu Sisouk , Julie Delon , Julien Tierny

Suppose we are given two metric spaces and a family of continuous transformations from one to the other. Given a probability distribution on each of these two spaces - namely the source and the target measures - the Wasserstein alignment…

Probability · Mathematics 2025-03-11 Soumik Pal , Bodhisattva Sen , Ting-Kam Leonard Wong

We study first-order optimality conditions for constrained optimization in the Wasserstein space, whereby one seeks to minimize a real-valued function over the space of probability measures endowed with the Wasserstein distance. Our…

Optimization and Control · Mathematics 2025-03-03 Nicolas Lanzetti , Saverio Bolognani , Florian Dörfler

Missing data imputation, where a model is trained on observed data to estimate unobserved values, is a fundamental problem in machine learning. In this paper, we rigorously formulate imputation model learning as a mean-squared error risk…

Machine Learning · Statistics 2026-05-14 Luke Shannon , Song Liu , Katarzyna Reluga

We consider synthesis and analysis of probability measures using the entropy-regularized Wasserstein-2 cost and its unbiased version, the Sinkhorn divergence. The synthesis problem consists of computing the barycenter, with respect to these…

Machine Learning · Statistics 2025-03-25 Brendan Mallery , James M. Murphy , Shuchin Aeron

In this paper, we study the problem of sampling from a distribution under the constraint of differential privacy (DP). Prior works measure the utility of DP sampling with density ratio-based measures such as KL divergence. However, such…

Machine Learning · Statistics 2026-05-12 Shokichi Takakura , Seng Pei Liew , Satoshi Hasegawa

Conditional distribution is a fundamental quantity for describing the relationship between a response and a predictor. We propose a Wasserstein generative approach to learning a conditional distribution. The proposed approach uses a…

Machine Learning · Computer Science 2021-12-21 Shiao Liu , Xingyu Zhou , Yuling Jiao , Jian Huang

Many machine learning problems can be expressed as the optimization of some cost functional over a parametric family of probability distributions. It is often beneficial to solve such optimization problems using natural gradient methods.…

Machine Learning · Statistics 2020-02-14 Michael Arbel , Arthur Gretton , Wuchen Li , Guido Montufar

Determinantal point processes (DPPs) have received significant attention as an elegant probabilistic model for discrete subset selection. Most prior work on DPP learning focuses on maximum likelihood estimation (MLE). While efficient and…

Machine Learning · Computer Science 2020-11-20 Lucas Anquetil , Mike Gartrell , Alain Rakotomamonjy , Ugo Tanielian , Clément Calauzènes

Stochastic approximation (SA) is a method for finding the root of an operator perturbed by noise. There is a rich literature establishing the asymptotic normality of rescaled SA iterates under fairly mild conditions. However, these…

Machine Learning · Statistics 2026-02-17 Shaan Ul Haque , Zedong Wang , Zixuan Zhang , Siva Theja Maguluri

We study the problem of robust distribution estimation under the Wasserstein distance, a popular discrepancy measure between probability distributions rooted in optimal transport (OT) theory. Given $n$ samples from an unknown distribution…

Machine Learning · Statistics 2024-09-25 Sloan Nietert , Rachel Cummings , Ziv Goldfeld