English
Related papers

Related papers: Minimax Statistical Estimation under Wasserstein C…

200 papers

Optimal transport and Wasserstein distances are flourishing in many scientific fields as a means for comparing and connecting random structures. Here we pioneer the use of an optimal transport distance between L\'{e}vy measures to solve a…

Statistics Theory · Mathematics 2023-09-18 Marta Catalano , Hugo Lavenant , Antonio Lijoi , Igor Prünster

A wide array of machine learning problems are formulated as the minimization of the expectation of a convex loss function on some parameter space. Since the probability distribution of the data of interest is usually unknown, it is is often…

Optimization and Control · Mathematics 2019-05-27 Emilie Chouzenoux , Henri Gérard , Jean-Christophe Pesquet

The problem of univariate mean change point detection and localization based on a sequence of $n$ independent observations with piecewise constant means has been intensively studied for more than half century, and serves as a blueprint for…

Statistics Theory · Mathematics 2019-06-07 Daren Wang , Yi Yu , Alessandro Rinaldo

Optimal transport is widely used to learn distributions, enforce distributional constraints, and model uncertainty. In applications, transport losses are often computed from samples through tractable representations, such as one-dimensional…

Optimization and Control · Mathematics 2026-05-28 Tam Le

The Wasserstein distance is an attractive tool for data analysis but statistical inference is hindered by the lack of distributional limits. To overcome this obstacle, for probability measures supported on finitely many points, we derive…

Methodology · Statistics 2017-04-27 Max Sommerfeld , Axel Munk

Problem definition: A key challenge in supervised learning is data scarcity, which can cause prediction models to overfit to the training data and perform poorly out of sample. A contemporary approach to combat overfitting is offered by…

Optimization and Control · Mathematics 2025-10-10 Reza Belbasi , Aras Selvi , Wolfram Wiesemann

This paper presents a unified approach based on Wasserstein distance to derive concentration bounds for empirical estimates for two broad classes of risk measures defined in the paper. The classes of risk measures introduced include as…

Statistics Theory · Mathematics 2022-05-11 Prashanth L. A. , Sanjay P. Bhat

For $\ell\colon \mathbb{R}^d \to [0,\infty)$ we consider the sequence of probability measures $\left(\mu_n\right)_{n \in \mathbb{N}}$, where $\mu_n$ is determined by a density that is proportional to $\exp(-n\ell)$. We allow for infinitely…

Probability · Mathematics 2023-12-11 Mareike Hasenpflug , Daniel Rudolf , Björn Sprungk

For Huber contamination on a known finite sample space, the unrestricted contaminating law is a probability vector on the support atoms, and domination over all measurable subsets reduces to atomwise inequalities. Placing a Dirichlet prior…

Methodology · Statistics 2026-05-27 Jaehoan Kim

A general lower bound is developed for the minimax risk when estimating an arbitrary functional. The bound is based on testing two composite hypotheses and is shown to be effective in estimating the nonsmooth functional…

Statistics Theory · Mathematics 2011-05-17 T. Tony Cai , Mark G. Low

In this paper, we explore a static setting for the assessment of risk in the context of mathematical finance and actuarial science that takes into account model uncertainty in the distribution of a possibly infinite-dimensional risk factor.…

Risk Management · Quantitative Finance 2024-08-13 Max Nendel , Alessandro Sgarabottolo

This paper presents a number of new findings about the canonical change point estimation problem. The first part studies the estimation of a change point on the real line in a simple stump model using the robust Huber estimating function…

Statistics Theory · Mathematics 2021-05-26 Debarghya Mukherjee , Moulinath Banerjee , Ya'acov Ritov

We propose a hybrid resampling method to approximate finitely supported Wasserstein barycenters on large-scale datasets, which can be combined with any exact solver. Nonasymptotic bounds on the expected error of the objective value as well…

Computation · Statistics 2021-05-28 Florian Heinemann , Axel Munk , Yoav Zemel

Wasserstein distributionally robust optimization (WDRO) attempts to learn a model that minimizes the local worst-case risk in the vicinity of the empirical data distribution defined by Wasserstein ball. While WDRO has received attention as…

Machine Learning · Statistics 2020-06-23 Yongchan Kwon , Wonyoung Kim , Joong-Ho Won , Myunghee Cho Paik

When propagating uncertainty in the data of differential equations, the probability laws describing the uncertainty are typically themselves subject to uncertainty. We present a sensitivity analysis of uncertainty propagation for…

Probability · Mathematics 2022-03-01 Oliver G. Ernst , Alois Pichler , Björn Sprungk

We study a model for adversarial classification based on distributionally robust chance constraints. We show that under Wasserstein ambiguity, the model aims to minimize the conditional value-at-risk of the distance to misclassification,…

Machine Learning · Computer Science 2021-11-05 Nam Ho-Nguyen , Stephen J. Wright

The distance and divergence of the probability measures play a central role in statistics, machine learning, and many other related fields. The Wasserstein distance has received much attention in recent years because of its distinctions…

Statistics Theory · Mathematics 2021-03-11 Qijun Tong , Kei Kobayashi

Conditional estimation given specific covariate values (i.e., local conditional estimation or functional estimation) is ubiquitously useful with applications in engineering, social and natural sciences. Existing data-driven non-parametric…

Machine Learning · Statistics 2020-10-13 Viet Anh Nguyen , Fan Zhang , Jose Blanchet , Erick Delage , Yinyu Ye

We derive quantitative bounds on the rate of convergence in $L^1$ Wasserstein distance of general M-estimators, with an almost sharp (up to a logarithmic term) behavior in the number of observations. We focus on situations where the…

Statistics Theory · Mathematics 2021-11-19 François Bachoc , Max Fathi

In this work, we study the weighted empirical risk minimization (weighted ERM) schema, in which an additional data-dependent weight function is incorporated when the empirical risk function is being minimized. We show that under a general…

Machine Learning · Computer Science 2025-01-07 Yikai Zhang , Jiahe Lin , Fengpei Li , Songzhu Zheng , Anant Raj , Anderson Schneider , Yuriy Nevmyvaka