English
Related papers

Related papers: Sharp concentration of uniform generalization erro…

200 papers

We investigate quantitative implications of the notion of log-concavity through a probabilistic interpretation. In particular, we derive concentration inequalities, moment and entropy bounds for random variables satisfying a precise degree…

Probability · Mathematics 2026-02-19 Arnaud Marsiglietti , James Melbourne

We consider the problem of exact recovery of a $k$-sparse binary vector from generalized linear measurements (such as logistic regression). We analyze the linear estimation algorithm (Plan, Vershynin, Yudovina, 2017), and also show…

Machine Learning · Statistics 2025-02-25 Arya Mazumdar , Neha Sangwan

New Vapnik and Chervonenkis type concentration inequalities are derived for the empirical distribution of an independent random sample. Focus is on the maximal deviation over classes of Borel sets within a low probability region. The…

Statistics Theory · Mathematics 2022-04-26 Stéphane Lhaut , Anne Sabourin , Johan Segers

A standard approach in pattern classification is to estimate the distributions of the label classes, and then to apply the Bayes classifier to the estimates of the distributions in order to classify unlabeled examples. As one might expect,…

Machine Learning · Computer Science 2007-05-23 Nick Palmer , Paul W. Goldberg

We prove risk bounds for binary classification in high-dimensional settings when the sample size is allowed to be smaller than the dimensionality of the training set observations. In particular, we prove upper bounds for both 'compressive…

Statistics Theory · Mathematics 2017-09-29 Ata Kaban , Robert J. Durrant

Under the Kolmogorov--Smirnov metric, an upper bound on the rate of convergence to the Gaussian distribution is obtained for linear statistics of the matrix ensembles in the case of the Gaussian, Laguerre, and Jacobi weights. The main lemma…

Probability · Mathematics 2020-06-16 Sergey Berezin , Alexander I. Bufetov

The assumption that response and predictor belong to the same statistical unit may be violated in practice. Unbiased estimation and recovery of true label ordering based on unlabeled data are challenging tasks and have attracted increasing…

Methodology · Statistics 2022-06-24 Guanhua Fang , Ping Li

This paper provides data-dependent bounds on the expected error of the Gibbs algorithm in the overparameterized interpolation regime, where low training errors are also obtained for impossible data, such as random labels in classification.…

Machine Learning · Computer Science 2026-02-13 Andreas Maurer , Erfan Mirzaei , Massimiliano Pontil

Building on the inequalities for homogeneous tetrahedral polynomials in independent Gaussian variables due to R. Lata{\l}a we provide a concentration inequality for non-necessarily Lipschitz functions $f\colon \R^n \to \R$ with bounded…

Probability · Mathematics 2013-04-09 Radosław Adamczak , Paweł Wolff

Matrix concentration inequalities provide information about the probability that a random matrix is close to its expectation with respect to the $l_2$ operator norm. This paper uses semigroup methods to derive sharp nonlinear matrix…

Probability · Mathematics 2021-01-08 De Huang , Joel A. Tropp

Several researchers have experimentally shown that substantial improvements can be obtained in difficult pattern recognition problems by combining or integrating the outputs of multiple classifiers. This chapter provides an analytical…

Neural and Evolutionary Computing · Computer Science 2007-05-23 Kagan Tumer , Joydeep Ghosh

Error bounds have been studied for more than seventy years, beginning with the seminal result of Hoffman (1952) [{\it J. Res. Natl. Bur. Standards}, 49 (1952), 263--265], which establishes an upper bound for the distance from an arbitrary…

Optimization and Control · Mathematics 2026-05-25 Zhou Wei , Michel Thera , Jen-Chih Yao

We develop a unified approach to universality of local scaling limits for eigenvalues of random normal matrices, or equivalently for planar Coulomb gases at inverse temperature $\beta=2$. The approach is direct in that it does not rely on…

Probability · Mathematics 2025-11-25 Joakim Cronvall , Aron Wennman

Sup-norm curve estimation is a fundamental statistical problem and, in principle, a premise for the construction of confidence bands for infinite-dimensional parameters. In a Bayesian framework, the issue of whether the…

Methodology · Statistics 2016-03-22 Catia Scricciolo

Rigorous statistical methods, including parameter estimation with accompanying uncertainties, underpin the validity of scientific discovery, especially in the natural sciences. With increasingly complex data models such as deep learning…

Machine Learning · Computer Science 2026-02-18 Aurora Grefsrud , Nello Blaser , Trygve Buanes

For general ferromagnetic Ising models whose coupling matrix has bounded spectral radius, we show that the log-Sobolev constant satisfies a simple bound expressed only in terms of the susceptibility of the model. This bound implies very…

Probability · Mathematics 2024-04-25 Roland Bauerschmidt , Benoit Dagallier

We study weighted Tikhonov regularization for large-scale linear discrete ill-posed problems with random noise. Under a polynomial upper-bound assumption on the generalized eigenvalues of the discrete forward operator, we derive stochastic…

Numerical Analysis · Mathematics 2026-05-19 Duan-Peng Ling , Wenlong Zhang

We study in this paper lower bounds for the generalization error of models derived from multi-layer neural networks, in the regime where the size of the layers is commensurate with the number of samples in the training data. We show that…

Machine Learning · Statistics 2022-07-08 Inbar Seroussi , Ofer Zeitouni

Over-parameterized models can perfectly learn various types of data distributions, however, generalization error is usually lower for real data in comparison to artificial data. This suggests that the properties of data distributions have…

Machine Learning · Computer Science 2022-07-28 Martin Briesch , Dominik Sobania , Franz Rothlauf

We consider semi-supervised classification when part of the available data is unlabeled. These unlabeled data can be useful for the classification problem when we make an assumption relating the behavior of the regression function to that…

Statistics Theory · Mathematics 2007-06-13 Philippe Rigollet
‹ Prev 1 4 5 6 7 8 10 Next ›