English
Related papers

Related papers: How many moments does MMD compare?

200 papers

The Word Movers Distance (WMD) measures the semantic dissimilarity between two text documents by computing the cost of optimally moving all words of a source/query document to the most similar words of a target document. Computing WMD…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-04-27 Jesmin Jahan Tithi , Fabrizio Petrini

This paper introduces kdiff, a novel kernel-based measure for estimating distances between instances of time series, random fields and other forms of structured data. This measure is based on the idea of matching distributions that only…

Machine Learning · Statistics 2021-10-01 Srinjoy Das , Hrushikesh Mhaskar , Alexander Cloninger

The Maximum Mean Discrepancy (MMD) is a cornerstone statistic for nonparametric two-sample testing, but its test power is dictated entirely by the chosen kernel. Because any fixed kernel inherently fails to distinguish certain…

Machine Learning · Statistics 2026-05-11 Yijin Ni , Xiaoming Huo

Approximate Markov chain Monte Carlo (MCMC) offers the promise of more rapid sampling at the cost of more biased inference. Since standard MCMC diagnostics fail to detect these biases, researchers have developed computable Stein discrepancy…

Machine Learning · Statistics 2020-10-16 Jackson Gorham , Lester Mackey

Kernel mean embeddings have recently attracted the attention of the machine learning community. They map measures $\mu$ from some set $M$ to functions in a reproducing kernel Hilbert space (RKHS) with kernel $k$. The RKHS distance of two…

Machine Learning · Statistics 2019-12-18 Carl-Johann Simon-Gabriel , Bernhard Schölkopf

Kernel embeddings of distributions and the Maximum Mean Discrepancy (MMD), the resulting distance between distributions, are useful tools for fully nonparametric two-sample testing and learning on distributions. However, it is rarely that…

Machine Learning · Statistics 2017-11-07 Ho Chung Leon Law , Christopher Yau , Dino Sejdinovic

We propose a nonparametric two-sample test procedure based on Maximum Mean Discrepancy (MMD) for testing the hypothesis that two samples of functions have the same underlying distribution, using kernels defined on function spaces. This…

Statistics Theory · Mathematics 2020-10-20 George Wynne , Andrew B. Duncan

Kernel techniques are among the most popular and flexible approaches in data science allowing to represent probability measures without loss of information under mild conditions. The resulting mapping called mean embedding gives rise to a…

Machine Learning · Statistics 2024-11-27 Linda Chamakh , Zoltan Szabo

The approximation of a discrete probability distribution $\mathbf{t}$ by an $M$-type distribution $\mathbf{p}$ is considered. The approximation error is measured by the informational divergence $\mathbb{D}(\mathbf{t}\Vert\mathbf{p})$, which…

Information Theory · Computer Science 2016-07-28 Bernhard C. Geiger , Georg Böcherer

Biclustering algorithms partition data and covariates simultaneously, providing new insights in several domains, such as analyzing gene expression to discover new biological functions. This paper develops a new model-free biclustering…

Methodology · Statistics 2022-08-09 Marcos Matabuena , J. C Vidal , Oscar Hernan Madrid Padilla , Dino Sejdinovic

We characterize the reproducing kernel Hilbert spaces whose elements are $p$-integrable functions in terms of the boundedness of the integral operator whose kernel is the reproducing kernel. Moreover, for $p=2$ we show that the spectral…

Functional Analysis · Mathematics 2007-05-23 Claudio Carmeli , Ernesto De Vito , Alessandro Toigo

The Maximum Mean Discrepancy (MMD) has been the state-of-the-art nonparametric test for tackling the two-sample problem. Its statistic is given by the difference in expectations of the witness function, a real-valued function defined as a…

Machine Learning · Computer Science 2022-02-14 Jonas M. Kübler , Wittawat Jitkrittum , Bernhard Schölkopf , Krikamol Muandet

Let $\{X_n\}_{n\in\N}$ be a Markov chain on a measurable space $\X$ with transition kernel $P$ and let $V:\X\r[1,+\infty)$. The Markov kernel $P$ is here considered as a linear bounded operator on the weighted-supremum space $\cB_V$…

Probability · Mathematics 2013-12-06 Loïc Hervé , James Ledoux

In this article, we begin a systematic study of the boundedness and the nuclearity properties of multilinear periodic pseudo-differential operators and multilinear discrete pseudo-differential operators on $L^p$-spaces. First, we prove…

Functional Analysis · Mathematics 2019-07-24 Duván Cardona , Vishvesh Kumar

We study a new class of pseudo differential operators whose symbols satisfy the differential inequality with a mixture of homogeneities. On the other hand, by taking singular integral realization, it can be equivalently defined by kernels…

Functional Analysis · Mathematics 2023-07-04 Zipeng Wang

We deal with kernel theorems for modulation spaces. We completely characterize the continuity of a linear operator on the modulation spaces $M^p$ for every $1\leq p\leq\infty$, by the membership of its kernel to (mixed) modulation spaces.…

Functional Analysis · Mathematics 2018-03-23 Elena Cordero , Fabio Nicola

The purpose of this paper is to study the $L^p$ boundedness of operators of the form \[ f\mapsto \psi(x) \int f(\gamma_t(x))K(t)\: dt, \] where $\gamma_t(x)$ is a $C^\infty$ function defined on a neighborhood of the origin in $(t,x)\in…

Classical Analysis and ODEs · Mathematics 2013-08-01 Elias M. Stein , Brian Street

Studying the stability of partially observed Markov decision processes (POMDPs) with respect to perturbations in either transition or observation kernels is a significant problem. While asymptotic robustness/stability results as approximate…

Optimization and Control · Mathematics 2025-09-15 Yunus Emre Demirci , Ali Devran Kara , Serdar Yüksel

Much recent work in bioinformatics has focused on the inference of various types of biological networks, representing gene regulation, metabolic processes, protein-protein interactions, etc. A common setting involves inferring network edges…

Quantitative Methods · Quantitative Biology 2007-05-23 Jean-Philippe Vert , Jian Qiu , William Stafford Noble

We consider the problem of simultaneously learning to linearly combine a very large number of kernels and learn a good predictor based on the learnt kernel. When the number of kernels $d$ to be combined is very large, multiple kernel…

Machine Learning · Computer Science 2015-03-20 Arash Afkanpour , András György , Csaba Szepesvári , Michael Bowling