English
Related papers

Related papers: Hilbert Curve Projection Distance for Distribution…

200 papers

We introduce and apply Hilbert's projective metric in the context of quantum information theory. The metric is induced by convex cones such as the sets of positive, separable or PPT operators. It provides bounds on measures for statistical…

Mathematical Physics · Physics 2011-08-16 David Reeb , Michael J. Kastoryano , Michael M. Wolf

Clustering is an important exploratory data analysis technique to group objects based on their similarity. The widely used $K$-means clustering method relies on some notion of distance to partition data into a fewer number of groups. In the…

Machine Learning · Statistics 2022-10-14 Yubo Zhuang , Xiaohui Chen , Yun Yang

We consider Sobolev-type distances on probability measures over separable Hilbert spaces involving the Schatten-$p$ norms, which include as special cases a distance first introduced by Bourguin and Campese (2020) when $p=2$, and a distance…

Probability · Mathematics 2026-02-03 Federico Bassetti , Solesne Bourguin , Simon Campese , Giovanni Peccati

Sliced-Wasserstein distance (SW) and its variant, Max Sliced-Wasserstein distance (Max-SW), have been used widely in the recent years due to their fast computation and scalability even when the probability measures lie in a very high…

Machine Learning · Statistics 2020-10-06 Khai Nguyen , Nhat Ho , Tung Pham , Hung Bui

We provide upper bounds of the expected Wasserstein distance between a probability measure and its empirical version, generalizing recent results for finite dimensional Euclidean spaces and bounded functional spaces. Such a generalization…

Statistics Theory · Mathematics 2020-01-29 Jing Lei

We introduce sliced optimal transport dataset distance (s-OTDD), a model-agnostic, embedding-agnostic approach for dataset comparison that requires no training, is robust to variations in the number of classes, and can handle disjoint label…

Machine Learning · Computer Science 2025-05-16 Khai Nguyen , Hai Nguyen , Tuan Pham , Nhat Ho

Gilbert proposed an algorithm for bounding the distance between a given point and a convex set. In this article we apply the Gilbert's algorithm to get an upper bound on the Hilbert-Schmidt distance between a given state and the set of…

Quantum Physics · Physics 2020-07-30 Palash Pandya , Omer Sakarya , Marcin Wieśniak

Optimal transport and its related problems, including optimal partial transport, have proven to be valuable tools in machine learning for computing meaningful distances between probability or positive measures. This success has led to a…

Machine Learning · Computer Science 2023-07-26 Xinran Liu , Yikun Bai , Huy Tran , Zhanqi Zhu , Matthew Thorpe , Soheil Kolouri

Squared Wasserstein distance is a frequently used tool to measure discrepancy between probability distributions. This distance is typically computed between empirical measures of size $n$ from two underlying random samples. Unfortunately,…

Machine Learning · Statistics 2026-05-20 Peter Matthew Jacobs , Jeff M. Phillips

We introduce sparse random projection, an important dimension-reduction tool from machine learning, for the estimation of discrete-choice models with high-dimensional choice sets. Initially, high-dimensional data are compressed into a…

Machine Learning · Statistics 2016-04-21 Khai X. Chiong , Matthew Shum

The maximum mean discrepancy and Wasserstein distance are popular distance measures between distributions and play important roles in many machine learning problems such as metric learning, generative modeling, domain adaption, and…

Machine Learning · Computer Science 2025-01-22 Dong Qiao , Jicong Fan

This article provides an overview on the statistical modeling of complex data as increasingly encountered in modern data analysis. It is argued that such data can often be described as elements of a metric space that satisfies certain…

Methodology · Statistics 2024-02-28 Paromita Dubey , Yaqing Chen , Hans-Georg Müller

The Earth Mover's Distance (EMD) is a state-of-the art metric for comparing discrete probability distributions, but its high distinguishability comes at a high cost in computational complexity. Even though linear-complexity approximation…

Machine Learning · Computer Science 2019-05-29 Kubilay Atasu , Thomas Mittelholzer

This paper gives a self-contained introduction to the Hilbert projective metric $\mathcal{H}$ and its fundamental properties, with a particular focus on the space of probability measures. We start by defining the Hilbert pseudo-metric on…

Probability · Mathematics 2024-11-13 Samuel N. Cohen , Eliana Fausti

We formulate and solve a regression problem with time-stamped distributional data. Distributions are considered as points in the Wasserstein space of probability measures, metrized by the 2-Wasserstein metric, and may represent images,…

Systems and Control · Electrical Eng. & Systems 2021-06-30 Amirhossein Karimi , Tryphon T. Georgiou

We obtain an estimate for the expected subspace robust Wasserstein distance between any probability measure on the unit ball of a separable Hilbert space, and its empirical distribution from $n$ i.i.d. samples.

Probability · Mathematics 2025-12-05 Dakshesh Vasan

The optimal transport (OT) problem has gained significant traction in modern machine learning for its ability to: (1) provide versatile metrics, such as Wasserstein distances and their variants, and (2) determine optimal couplings between…

Machine Learning · Computer Science 2024-10-18 Xinran Liu , Rocío Díaz Martín , Yikun Bai , Ashkan Shahbazi , Matthew Thorpe , Akram Aldroubi , Soheil Kolouri

Discrepancy measures between probability distributions, often termed statistical distances, are ubiquitous in probability theory, statistics and machine learning. To combat the curse of dimensionality when estimating these distances from…

Statistics Theory · Mathematics 2021-12-21 Sloan Nietert , Ziv Goldfeld , Kengo Kato

Optimal transport distances (OT) have been widely used in recent work in Machine Learning as ways to compare probability distributions. These are costly to compute when the data lives in high dimension. Recent work by Paty et al., 2019,…

Machine Learning · Computer Science 2021-11-10 Patric M. Fulop , Vincent Danos

Testing for the equality of two high-dimensional distributions is a challenging problem, and this becomes even more challenging when the sample size is small. Over the last few decades, several graph-based two-sample tests have been…

Methodology · Statistics 2019-11-22 Soham Sarkar , Rahul Biswas , Anil K. Ghosh
‹ Prev 1 4 5 6 7 8 10 Next ›