English
Related papers

Related papers: Uniform Convergence Beyond Glivenko-Cantelli

200 papers

Heisenberg's uncertainty principle was originally posed for the limit of the accuracy of simultaneous measurement of non-commuting observables as stating that canonically conjugate observables can be measured simultaneously only with the…

Quantum Physics · Physics 2020-03-12 Masanao Ozawa

This paper addresses the statistical problem of estimating the infinite-norm deviation from the empirical mean to the distribution mean for high-dimensional distributions on $\{0,1\}^d$, potentially with $d=\infty$. Unlike traditional…

Statistics Theory · Mathematics 2024-02-21 Moïse Blanchard , Václav Voráček

This article considers the parametric estimation of $Pr(X<Y<Z)$ and its generalizations based on several well-known one-parameter and two-parameter continuous distributions. It is shown that for some one-parameter distributions and when…

Statistics Theory · Mathematics 2023-01-25 Tau Raphael Rasethuntsa

Creating accurate meta-embeddings from pre-trained source embeddings has received attention lately. Methods based on global and locally-linear transformation and concatenation have shown to produce accurate meta-embeddings. In this paper,…

Computation and Language · Computer Science 2018-04-17 Joshua Coates , Danushka Bollegala

It is common to assume in empirical research that observables and unobservables are additively separable, especially, when the former are endogenous. This is done because it is widely recognized that identification and estimation challenges…

Statistics Theory · Mathematics 2021-04-02 Andrii Babii , Jean-Pierre Florens

Ensembles are a straightforward, remarkably effective method for improving the accuracy,calibration, and robustness of models on classification tasks; yet, the reasons that underlie their success remain an active area of research. We build…

Machine Learning · Statistics 2022-06-22 Neha Gupta , Jamie Smith , Ben Adlam , Zelda Mariet

For many applications, an ensemble of base classifiers is an effective solution. The tuning of its parameters(number of classes, amount of data on which each classifier is to be trained on, etc.) requires G, the generalization error of a…

We establish a central limit theorem for the unnormalized linear statistic of the Gaussian Unitary Ensemble under optimal conditions: the linear statistics converges if and only if the expression for the limiting variance is finite.

Probability · Mathematics 2015-10-14 Phil Kopel

Consider informative selection of a sample from a finite population. Responses are realized as independent and identically distributed (i.i.d.) random variables with a probability density function (p.d.f.) f, referred to as the…

Statistics Theory · Mathematics 2012-11-26 Daniel Bonnéry , F. Jay Breidt , François Coquet

Randomness is ubiquitous in many applications across data science and machine learning. Remarkably, systems composed of random components often display emergent global behaviors that appear deterministic, manifesting a transition from…

Machine Learning · Computer Science 2025-05-16 Luca Muscarnera , Luigi Loreti , Giovanni Todeschini , Alessio Fumagalli , Francesco Regazzoni

Statistical learning theory chiefly studies restricted hypothesis classes, particularly those with finite Vapnik-Chervonenkis (VC) dimension. The fundamental quantity of interest is the sample complexity: the number of samples required to…

Machine Learning · Computer Science 2008-07-10 David Soloveichik

In this paper, we develop a computational approach for estimating the mean value of a quantity in the presence of uncertainty. We demonstrate that, under some mild assumptions, the upper and lower bounds of the mean value are efficiently…

Statistics Theory · Mathematics 2013-11-05 Xinjia Chen

We describe various moment-based ensemble interpretation models for the construction of probabilistic temperature forecasts from ensembles. We apply the methods to one year of medium range ensemble forecasts and perform in and out of sample…

Atmospheric and Oceanic Physics · Physics 2007-05-23 Stephen Jewson

New Vapnik and Chervonenkis type concentration inequalities are derived for the empirical distribution of an independent random sample. Focus is on the maximal deviation over classes of Borel sets within a low probability region. The…

Statistics Theory · Mathematics 2022-04-26 Stéphane Lhaut , Anne Sabourin , Johan Segers

Solomonoff's uncomputable universal prediction scheme $\xi$ allows to predict the next symbol $x_k$ of a sequence $x_1...x_{k-1}$ for any Turing computable, but otherwise unknown, probabilistic environment $\mu$. This scheme will be…

Machine Learning · Computer Science 2007-05-23 Marcus Hutter

We investigate a generalized empirical likelihood approach in a two-group setting where the constraints on parameters have a form of U-statistics. In this situation, the summands that consist of the constraints for the empirical likelihood…

Methodology · Statistics 2015-05-04 Jihnhee Yu , Luge Yang , Albert Vexler , Alan D. Hutson

The consistency of a learning method is usually established under the assumption that the observations are a realization of an independent and identically distributed (i.i.d.) or mixing process. Yet, kernel methods such as support vector…

Machine Learning · Computer Science 2024-06-11 Pierre-François Massiani , Sebastian Trimpe , Friedrich Solowjow

We extend the study of \emph{melonic} quartic tensor models to models with arbitrary quartic interactions. This extension requires a new version of the loop vertex expansion using several species of intermediate fields and iterated…

High Energy Physics - Theory · Physics 2017-06-26 Thibault Delepouve , Razvan Gurau , Vincent Rivasseau

We develop large sample theory for merged data from multiple sources. Main statistical issues treated in this paper are (1) the same unit potentially appears in multiple datasets from overlapping data sources, (2) duplicated items are not…

Statistics Theory · Mathematics 2018-05-22 Takumi Saegusa

Recent advances in center-based clustering continue to improve upon the drawbacks of Lloyd's celebrated $k$-means algorithm over $60$ years after its introduction. Various methods seek to address poor local minima, sensitivity to outliers,…

Machine Learning · Statistics 2021-10-28 Debolina Paul , Saptarshi Chakraborty , Swagatam Das , Jason Xu