English
Related papers

Related papers: The Dantzig selector and sparsity oracle inequalit…

200 papers

We point out some pitfalls related to the concept of an oracle property as used in Fan and Li (2001, 2002, 2004) which are reminiscent of the well-known pitfalls related to Hodges' estimator. The oracle property is often a consequence of…

Statistics Theory · Mathematics 2007-11-08 Hannes Leeb , Benedikt M. Poetscher

Let $\mathbf{A}$ be an $n\times n$-matrix over $\mathbb{F}_2$ whose every entry equals $1$ with probability $d/n$ independently for a fixed $d>0$. Draw a vector $\mathbf{y}$ randomly from the column space of $\mathbf{A}$. It is a simple…

Combinatorics · Mathematics 2023-09-08 Amin Coja-Oghlan , Oliver Cooley , Mihyun Kang , Joon Lee , Jean Bernoulli Ravelomanana

We study various constraints and conditions on the true coefficient vector and on the design matrix to establish non-asymptotic oracle inequalities for the prediction error, estimation accuracy and variable selection for the Lasso estimator…

Statistics Theory · Mathematics 2018-06-15 Niharika Gauraha

We study discrete random variants of the Carleson maximal operator. Intriguingly, these questions remain subtle and difficult, even in this setting. Let $\{X_m\}$ be an independent sequence of $\{0,1\}$ random variables with expectations \[…

Classical Analysis and ODEs · Mathematics 2016-09-29 Ben Krause , Michael T. Lacey

Let X_1,...., X_n be a collection of iid discrete random variables, and Y_1,..., Y_m a set of noisy observations of such variables. Assume each observation Y_a to be a random function of some a random subset of the X_i's, and consider the…

Information Theory · Computer Science 2007-09-04 Andrea Montanari

Popular sparse estimation methods based on $\ell_1$-relaxation, such as the Lasso and the Dantzig selector, require the knowledge of the variance of the noise in order to properly tune the regularization parameter. This constitutes a major…

Machine Learning · Statistics 2013-04-17 Arnak S. Dalalyan , Mohamed Hebiri , Katia Méziani , Joseph Salmon

We derive estimates for the largest and smallest singular values of sparse rectangular $N\times n$ random matrices, assuming $\lim_{N,n\to\infty}\frac nN=y\in(0,1)$. We consider a model with sparsity parameter $p_N$ such that $Np_N\sim…

Probability · Mathematics 2022-11-29 F. Götze , A. Tikhomirov

We consider a Gaussian sequence space model $X_{\lambda}=f_{\lambda} + \xi_{\lambda},$ where $\xi $ has a diagonal covariance matrix $\Sigma=\diag(\sigma_\lambda ^2)$. We consider the situation where the parameter vector $(f_{\lambda})$ is…

Statistics Theory · Mathematics 2013-12-23 Laurent Cavalier , Markus Reiß

We give oracle inequalities on procedures which combines quantization and variable selection via a weighted Lasso $k$-means type algorithm. The results are derived for a general family of weights, which can be tuned to size the influence of…

Statistics Theory · Mathematics 2016-07-07 Clément Levrard

Consider the problem of drawing random variates $(X_1,\ldots,X_n)$ from a distribution where the marginal of each $X_i$ is specified, as well as the correlation between every pair $X_i$ and $X_j$. For given marginals, the…

Probability · Mathematics 2016-12-30 Mark Huber , Nevena Maric

In this paper, we investigate the invertibility of sparse symmetric matrices. We show that for an $n\times n$ sparse symmetric random matrix $A$ with $A_{ij} = \delta_{ij} \xi_{ij}$ is invertible with high probability. Here, $\delta_{ij}$s,…

Probability · Mathematics 2018-04-26 Feng Wei

Let $R(n) = \sum_{a+b=n} \Lambda(a)\Lambda(b)$, where $\Lambda(\cdot)$ is the von Mangoldt function. The function $R(n)$ is often studied in connection with Goldbach's conjecture. On the Riemann hypothesis (RH) it is known that $\sum_{n\leq…

Number Theory · Mathematics 2020-06-29 Michael J. Mossinghoff , Timothy S. Trudgian

We study the maximum likelihood estimator of density of $n$ independent observations, under the assumption that it is well approximated by a mixture with a large number of components. The main focus is on statistical properties with respect…

Statistics Theory · Mathematics 2017-01-19 Arnak S. Dalalyan , Mehdi Sebbar

Let $(X_i, \mathcal{F}_i)_{i\geq1}$ be a martingale difference sequence in a smooth Banach space. Let $S_n=\sum_{i=1}^nX_i, n\geq 1,$ be the partial sums of $(X_i, \mathcal{F}_i)_{i\geq 1}$. We give upper bounds on the quantity…

Probability · Mathematics 2019-09-13 Xiequan Fan , Davide Giraudo

Inference and prediction under the sparsity assumption have been a hot research topic in recent years. However, in practice, the sparsity assumption is difficult to test, and more importantly can usually be violated. In this paper, to study…

Statistics Theory · Mathematics 2022-10-18 Yanmei Shi , Zhiruo Li , Qi Zhang

This paper is concerned with high-dimensional panel data models where the number of regressors can be much larger than the sample size. Under the assumption that the true parameter vector is sparse we propose a panel-Lasso estimator and…

Statistics Theory · Mathematics 2014-02-14 Anders Bredahl Kock

Following S\"odergren, we consider a collection of random variables on the space $X_n$ of unimodular lattices in dimension $n$: Normalizations of the angles between the $N = N(n)$ shortest vectors in a random unimodular lattice, and the…

Number Theory · Mathematics 2022-06-15 Kristian Holm

Transductive methods are useful in prediction problems when the training dataset is composed of a large number of unlabeled observations and a smaller number of labeled observations. In this paper, we propose an approach for developing…

Statistics Theory · Mathematics 2010-06-16 Pierre Alquier , Mohamed Hebiri

Sparse linear regression is one of the most basic questions in machine learning and statistics. Here, we are given as input a design matrix $X \in \mathbb{R}^{N \times d}$ and measurements or labels ${y} \in \mathbb{R}^N$ where ${y} = {X}…

Machine Learning · Computer Science 2025-11-11 Gautam Chandrasekaran , Raghu Meka , Konstantinos Stavropoulos

We consider a finite mixture of Gaussian regression model for high- dimensional data, where the number of covariates may be much larger than the sample size. We propose to estimate the unknown conditional mixture density by a maximum…

Statistics Theory · Mathematics 2014-09-05 Emilie Devijver