English
Related papers

Related papers: Nearly optimal central limit theorem and bootstrap…

200 papers

Due to their importance in both data analysis and numerical algorithms, low rank approximations have recently been widely studied. They enable the handling of very large matrices. Tight error bounds for the computationally efficient…

Numerical Analysis · Mathematics 2023-04-06 Frank de Hoog , Markus Hegland

In the context of principal components analysis (PCA), the bootstrap is commonly applied to solve a variety of inference problems, such as constructing confidence intervals for the eigenvalues of the population covariance matrix $\Sigma$.…

Statistics Theory · Mathematics 2022-02-17 Junwen Yao , Miles E. Lopes

A multidimensional version of the results of Koml\'os, Major and Tusn\'ady for sums of independent random vectors with finite exponential moments is obtained in the particular case where the summands have smooth distributions which are…

Probability · Mathematics 2014-02-07 F. Götze , A. Yu. Zaitsev

In the literature of high-dimensional central limit theorems, there is a gap between results for general limiting correlation matrix $\Sigma$ and the strongly non-degenerate case. For the general case where $\Sigma$ may be degenerate, under…

Probability · Mathematics 2023-05-30 Xiao Fang , Yuta Koike , Song-Hao Liu , Yi-Kun Zhao

Let $X=C+\mathrm{E}$ with a deterministic matrix $C\in\R^{M\times M}$ and $\mathrm{E}$ some centered Gaussian $M\times M$-matrix whose entries are independent with variance $\sigma^2$. In the present work, the accuracy of reduced-rank…

Probability · Mathematics 2012-05-08 Angelika Rohde

We study the distribution of a fully connected neural network with random Gaussian weights and biases in which the hidden layer widths are proportional to a large constant $n$. Under mild assumptions on the non-linearity, we obtain…

Machine Learning · Computer Science 2024-06-18 Stefano Favaro , Boris Hanin , Domenico Marinucci , Ivan Nourdin , Giovanni Peccati

This work considers the problem of estimating the distance between two covariance matrices directly from the data. Particularly, we are interested in the family of distances that can be expressed as sums of traces of functions that are…

Machine Learning · Computer Science 2024-09-19 Roberto Pereira , Xavier Mestre , Davig Gregoratti

Let $X$ be a $d$-dimensional random vector and $X_\theta$ its projection onto the span of a set of orthonormal vectors $\{\theta_1,...,\theta_k\}$. Conditions on the distribution of $X$ are given such that if $\theta$ is chosen according to…

Probability · Mathematics 2011-02-16 Elizabeth Meckes

In the first part of this paper we study a best approximation of a vector in Euclidean space R^n with respect to a closed semi-algebraic set C and a given semi-algebraic norm. Assuming that the given norm and its dual norm are…

Algebraic Geometry · Mathematics 2013-11-08 Shmuel Friedland , Malgorzata Stawiska

We obtain bounds to quantify the distributional approximation in the delta method for vector statistics (the sample mean of $n$ independent random vectors) for normal and non-normal limits, measured using smooth test functions. For normal…

Statistics Theory · Mathematics 2023-05-11 Robert E. Gaunt , Heather Sutcliffe

We study the total $\alpha$-powered length of the rooted edges in a random minimal directed spanning tree - first introduced in Bhatt and Roy (2004) - on a Poisson process with intensity $s \ge 1$ on the unit cube $[0,1]^d$ for $d \ge 3$.…

Probability · Mathematics 2022-12-06 Chinmoy Bhattacharjee

We study a random partial covering model on the $(d-1)$-dimensional unit sphere, where $N$ spherical caps are placed independently and uniformly at random, each covering a surface fraction of $1/N$. This model provides a continuous…

Probability · Mathematics 2026-04-10 Steven Hoehner , Christoph Thäle

Nearest neighbor cells in $R^d,d\in\mathbb{N}$, are used to define coefficients of divergence ($\phi$-divergences) between continuous multivariate samples. For large sample sizes, such distances are shown to be asymptotically normal with a…

Probability · Mathematics 2009-03-06 Yu. Baryshnikov , Mathew D. Penrose , J. E. Yukich

We consider the problem of testing the mean of high-dimensional data when the dimension may grow without explicit rate restrictions relative to the sample size. The proposed procedure is based on the statistic V_n = n||Xn||^2, which avoids…

Statistics Theory · Mathematics 2026-05-18 Dietmar Ferger

This paper considers a new bootstrap procedure to estimate the distribution of high-dimensional $\ell_p$-statistics, i.e. the $\ell_p$-norms of the sum of $n$ independent $d$-dimensional random vectors with $d \gg n$ and $p \in [1,…

Statistics Theory · Mathematics 2020-08-18 Alexander Giessing , Jianqing Fan

For time series with long-range temporal dependence, inference for covariance and precision matrices is non-trivial. We propose a Berry-Esseen type Gaussian approximation result that gives a finite-sample bound for the Kolmogorov distance…

Statistics Theory · Mathematics 2026-04-20 Percy S. Zhai , Mladen Kolar , Wei Biao Wu

We study the scaling limit of essentially simple triangulations on the torus. We consider, for every $n\geq 1$, a uniformly random triangulation $G_n$ over the set of (appropriately rooted) essentially simple triangulations on the torus…

Discrete Mathematics · Computer Science 2019-05-07 Vincent Beffara , Cong Bang Huynh , Benjamin Lévêque

Central limit theorems (CLTs) for high-dimensional random vectors with dimension possibly growing with the sample size have received a lot of attention in the recent times. Chernozhukov et al. (2017) proved a Berry--Esseen type result for…

Statistics Theory · Mathematics 2019-06-26 Arun Kumar Kuchibhotla , Somabha Mukherjee , Debapratim Banerjee

Given a probability distribution in R^n with general (non-white) covariance, a classical estimator of the covariance matrix is the sample covariance matrix obtained from a sample of N independent points. What is the optimal sample size N =…

Probability · Mathematics 2014-05-21 Roman Vershynin

Generalized Linear Model (or GLM) extends the ordinary linear regression by linking the mean of the response variable to covariates through appropriate link functions. GLM is widely used in the analysis of datasets arising from diverse…

Methodology · Statistics 2026-04-28 Mayukh Choudhury , Debraj Das