English
Related papers

Related papers: Sharp Rates of MMD Empirical Estimation with Power…

200 papers

Marginal maximum likelihood estimation (MMLE) in item response theory (IRT) is highly sensitive to aberrant responses, such as careless answering and random guessing, which can reduce estimation accuracy. To address this issue, this study…

Methodology · Statistics 2025-02-18 Yuki Itaya , Kenichi Hayashi

Estimating the rate of convergence of the empirical measure of an i.i.d. sample to the reference measure is a classical problem in probability theory. Extending recent results of Ambrosio, Stra and Trevisan on 2-dimensional manifolds, in…

Probability · Mathematics 2024-02-20 Bence Borda

We propose a nonparametric two-sample test procedure based on Maximum Mean Discrepancy (MMD) for testing the hypothesis that two samples of functions have the same underlying distribution, using kernels defined on function spaces. This…

Statistics Theory · Mathematics 2020-10-20 George Wynne , Andrew B. Duncan

Maximum mean discrepancy (MMD) has been widely employed to measure the distance between probability distributions. In this paper, we propose using MMD to solve continuous multi-objective optimization problems (MOPs). For solving MOPs, a…

Machine Learning · Computer Science 2025-05-21 Hao Wang , Chenyu Shi , Angel E. Rodriguez-Fernandez , Oliver Schütze

We consider a symmetric mixture of linear regressions with random samples from the pairwise comparison design, which can be seen as a noisy version of a type of Euclidean distance geometry problem. We analyze the expectation-maximization…

Statistics Theory · Mathematics 2023-06-23 Abhishek Dhawan , Cheng Mao , Ashwin Pananjady

We characterize the asymptotic performance of nonparametric one- and two-sample testing. The exponential decay rate or error exponent of the type-II error probability is used as the asymptotic performance metric, and an optimal test…

Information Theory · Computer Science 2021-02-08 Shengyu Zhu , Biao Chen , Zhitang Chen , Pengfei Yang

The distribution closeness testing (DCT) assesses whether the distance between a distribution pair is at least $\epsilon$-far. Existing DCT methods mainly measure discrepancies between a distribution pair defined on discrete one-dimensional…

Machine Learning · Computer Science 2025-10-10 Zhijian Zhou , Liuhua Peng , Xunye Tian , Feng Liu

Let $\mu$ be an Ahlfors-David probability measure on $\mathbb{R}^q$, namely, there exist some constants $s_0>0$ and $\epsilon_0,C_1,C_2>0$ such that \[ C_1\epsilon^{s_0}\leq\mu(B(x,\epsilon))\leq…

Metric Geometry · Mathematics 2018-02-27 Sanguo Zhu

Let $(X,d,\mu)$ be a $RCD^\ast(K, N)$ space with $K\in \mathbb{R}$ and $N\in [1,\infty]$. For $N\in [1,\infty)$, we derive the upper and lower bounds of the heat kernel on $(X,d,\mu)$ by applying the parabolic Harnack inequality and the…

Metric Geometry · Mathematics 2015-12-02 Renjin Jiang , Huaiqian Li , Huichun Zhang

We introduce kernel thinning, a new procedure for compressing a distribution $\mathbb{P}$ more effectively than i.i.d. sampling or standard thinning. Given a suitable reproducing kernel $\mathbf{k}_{\star}$ and $O(n^2)$ time, kernel…

Machine Learning · Statistics 2024-05-14 Raaz Dwivedi , Lester Mackey

Distance metrics are central to machine learning, yet distances between ensembles of quantum states remain poorly understood due to fundamental quantum measurement constraints. We introduce a hierarchy of integral probability metrics,…

Quantum Physics · Physics 2026-01-30 Jian Yao , Pengtao Li , Xiaohui Chen , Quntao Zhuang

Nonparametric two sample testing is a decision theoretic problem that involves identifying differences between two random variables without making parametric assumptions about their underlying distributions. We refer to the most common…

Statistics Theory · Mathematics 2015-08-05 Aaditya Ramdas , Sashank J. Reddi , Barnabas Poczos , Aarti Singh , Larry Wasserman

Distances between probability distributions are a key component of many statistical machine learning tasks, from two-sample testing to generative modeling, among others. We introduce a novel distance between measures that compares them…

Machine Learning · Statistics 2025-07-09 Arturo Castellanos , Anna Korba , Pavlo Mozharovskyi , Hicham Janati

Given an isotropic probability measure $\mu$ on ${\mathbb R}^d$ with ${\rm d}\mu \left( x \right) = {\left( {\varrho \left( x \right)} \right)^{ - \alpha }}{\rm d}x$, where $\alpha > d + 1$ and $\varrho :{{\mathbb R}^d} \to \left( {0, +…

Probability · Mathematics 2021-03-08 Huynh Khanh , Filippo Santambrogio , Doan Thai Son

An important feature of kernel mean embeddings (KME) is that the rate of convergence of the empirical KME to the true distribution KME can be bounded independently of the dimension of the space, properties of the distribution and smoothness…

Statistics Theory · Mathematics 2025-04-17 Geoffrey Wolfer , Pierre Alquier

We propose conditional flows of the maximum mean discrepancy (MMD) with the negative distance kernel for posterior sampling and conditional generative modeling. This MMD, which is also known as energy distance, has several advantageous…

We consider the problem of approximating an analytic function on a compact interval from its values at $M+1$ distinct points. When the points are equispaced, a recent result (the so-called impossibility theorem) has shown that the best…

Numerical Analysis · Mathematics 2018-04-09 Ben Adcock , Rodrigo Platte , Alexei Shadrin

Negative distance kernels $K(x,y) := - \|x-y\|$ were used in the definition of maximum mean discrepancies (MMDs) in statistics and lead to favorable numerical results in various applications. In particular, so-called slicing techniques for…

Machine Learning · Statistics 2025-10-23 Nicolaj Rux , Michael Quellmalz , Gabriele Steidl

Let $\Xi$ be the adjacency matrix of an Erd\H{o}s-R\'enyi graph on $n$ vertices and with parameter $p$ and consider $A$ a $n\times n$ centered random symmetric matrix with bounded i.i.d. entries above the diagonal. When the mean degree $np$…

Probability · Mathematics 2024-01-23 Fanny Augeri

Maximum Mean Discrepancy (MMD) is a widely used concept in machine learning research which has gained popularity in recent years as a highly effective tool for comparing (finite-dimensional) distributions. Since it is designed as a…

Machine Learning · Statistics 2025-06-03 Andrew Alden , Blanka Horvath , Zacharia Issa