English
Related papers

Related papers: Thinning a Wishart Random Matrix

200 papers

For a given $p\times n$ data matrix $\textbf{X}_n$ with i.i.d. centered entries and a population covariance matrix $\bf{\Sigma}$, the corresponding sample precision matrix $\hat{\bf\Sigma}^{-1}$ is defined as the inverse of the sample…

Statistics Theory · Mathematics 2022-12-21 Nina Dörnemann , Holger Dette

The distributions of the largest and the smallest eigenvalues of a $p$-variate sample covariance matrix $S$ are of great importance in statistics. Focusing on the null case where $nS$ follows the standard Wishart distribution $W_p(I,n)$, we…

Statistics Theory · Mathematics 2012-03-06 Zongming Ma

The asymptotic normality for a large family of eigenvalue statistics of a general sample covariance matrix is derived under the ultra-high dimensional setting, that is, when the dimension to sample size ratio $p/n \to \infty$. Based on this…

Methodology · Statistics 2021-09-15 Jiaxin Qiu , Zeng Li , Jianfeng Yao

There are many real-world classification problems wherein the issue of data imbalance (the case when a data set contains substantially more samples for one/many classes than the rest) is unavoidable. While under-sampling the problematic…

Computer Vision and Pattern Recognition · Computer Science 2018-01-09 John McKay , Isaac Gerg , Vishal Monga

Consider a standard white Wishart matrix with parameters $n$ and $p$. Motivated by applications in high-dimensional statistics and signal processing, we perform asymptotic analysis on the maxima and minima of the eigenvalues of all the $m…

Statistics Theory · Mathematics 2019-05-22 T. Tony Cai , Tiefeng Jiang , Xiaoou Li

In the statistical inference for long range dependent time series the shape of the limit distribution typically depends on unknown parameters. Therefore, we propose to use subsampling. We show the validity of subsampling for general…

Statistics Theory · Mathematics 2016-10-20 Annika Betken , Martin Wendler

Dataset distillation has emerged as a promising approach in deep learning, enabling efficient training with small synthetic datasets derived from larger real ones. Particularly, distribution matching-based distillation methods attract…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Wenxiao Deng , Wenbin Li , Tianyu Ding , Lei Wang , Hongguang Zhang , Kuihua Huang , Jing Huo , Yang Gao

Consider a linear model $Y=X\beta+z$, where $X=X_{n,p}$ and $z\sim N(0,I_n)$. The vector $\beta$ is unknown but is sparse in the sense that most of its coordinates are $0$. The main interest is to separate its nonzero coordinates from the…

Statistics Theory · Mathematics 2015-03-20 Zheng Tracy Ke , Jiashun Jin , Jianqing Fan

We introduce a pruning algorithm that provably sparsifies the parameters of a trained model in a way that approximately preserves the model's predictive accuracy. Our algorithm uses a small batch of input points to construct a data-informed…

Machine Learning · Computer Science 2021-03-16 Cenk Baykal , Lucas Liebenwein , Igor Gilitschenski , Dan Feldman , Daniela Rus

Subsampling is a computationally efficient and scalable method to draw inference in large data settings based on a subset of the data rather than needing to consider the whole dataset. When employing subsampling techniques, a crucial…

Methodology · Statistics 2025-10-08 Amalan Mahendran , Helen Thompson , James M. McGree

Starting with a set of weighted items, we want to create a generic sample of a certain size that we can later use to estimate the total weight of arbitrary subsets. For this purpose, we propose priority sampling which tested on Internet…

Data Structures and Algorithms · Computer Science 2007-05-23 Nick Duffield , Carsten Lund , Mikkel Thorup

Many modern tools in machine learning and signal processing, such as sparse dictionary learning, principal component analysis (PCA), non-negative matrix factorization (NMF), $K$-means clustering, etc., rely on the factorization of a matrix…

Machine Learning · Statistics 2015-04-10 Rémi Gribonval , Rodolphe Jenatton , Francis Bach , Martin Kleinsteuber , Matthias Seibert

An invariant ensemble of $N\times N$ random matrices can be characterised by a joint distribution for eigenvalues $P(\lambda_1,\cdots,\lambda_N)$. The study of the distribution of linear statistics, i.e. of quantities of the form…

Statistical Mechanics · Physics 2017-09-25 Aurélien Grabsch , Christophe Texier

Using a character expansion method, we calculate exactly the eigenvalue density of random matrices of the form M^\dagger M where M is a complex matrix drawn from a normalized distribution P(M) ~ exp(-\Tr(A M B M^\dagger) with A and B…

Mathematical Physics · Physics 2009-11-10 Steven H. Simon , Aris L. Moustakas

A Wishart matrix is said to be spiked when the underlying covariance matrix has a single eigenvalue $b$ different from unity. As $b$ increases through $b=2$, a gap forms from the largest eigenvalue to the rest of the spectrum, and with…

Mathematical Physics · Physics 2014-07-01 Peter J. Forrester

We propose a modification of linear discriminant analysis, referred to as compressive regularized discriminant analysis (CRDA), for analysis of high-dimensional datasets. CRDA is specially designed for feature elimination purpose and can be…

Methodology · Statistics 2018-04-12 Muhammad Naveed Tabassum , Esa Ollila

We consider large random matrices with a general slowly decaying correlation among its entries. We prove universality of the local eigenvalue statistics and optimal local laws for the resolvent away from the spectral edges, generalizing the…

Probability · Mathematics 2020-06-01 László Erdős , Torben Krüger , Dominik Schröder

We revisit the range sampling problem: the input is a set of points where each point is associated with a real-valued weight. The goal is to store them in a structure such that given a query range and an integer $k$, we can extract $k$…

Data Structures and Algorithms · Computer Science 2019-03-20 Peyman Afshani , Jeff M. Phillips

Matrix completion, i.e., the exact and provable recovery of a low-rank matrix from a small subset of its elements, is currently only known to be possible if the matrix satisfies a restrictive structural constraint---known as {\em…

Machine Learning · Statistics 2014-07-22 Yudong Chen , Srinadh Bhojanapalli , Sujay Sanghavi , Rachel Ward

This article introduces a general statistical modeling principle called "Density Sharpening" and applies it to the analysis of discrete count data. The underlying foundation is based on a new theory of nonparametric approximation and…

Methodology · Statistics 2021-08-24 Subhadeep Mukhopadhyay
‹ Prev 1 8 9 10 Next ›