中文
相关论文

相关论文: The sparsity and bias of the Lasso selection in hi…

200 篇论文

We study regression discontinuity designs in which many predetermined covariates, possibly much more than the number of observations, can be used to increase the precision of treatment effect estimates. We consider a two-step estimator…

计量经济学 · 经济学 2022-05-06 Alexander Kreiß , Christoph Rothe

Recent research has focused on $\ell_1$ penalized least squares (Lasso) estimators for high-dimensional linear regressions in which the number of covariates $p$ is considerably larger than the sample size $n$. However, few studies have…

统计理论 · 数学 2022-05-05 Yuefeng Han , Ruey S. Tsay

This paper explores the validity of the two-stage estimation procedure for sparse linear models in high-dimensional settings with possibly many endogenous regressors. In particular, the number of endogenous regressors in the main equation…

统计理论 · 数学 2013-09-18 Ying Zhu

In modern data analysis, sparse model selection becomes inevitable once the number of predictors variables is very high. It is well-known that model selection procedures like the Lasso or Boosting tend to overfit on real data. The…

机器学习 · 计算机科学 2022-02-11 Tino Werner

We study full Bayesian procedures for sparse linear regression when errors have a symmetric but otherwise unknown distribution. The unknown error distribution is endowed with a symmetrized Dirichlet process mixture of Gaussians. For the…

统计理论 · 数学 2019-03-26 Minwoo Chae , Lizhen Lin , David B. Dunson

We investigate the high-dimensional linear regression problem in the presence of noise correlated with Gaussian covariates. This correlation, known as endogeneity in regression models, often arises from unobserved variables and other…

统计理论 · 数学 2023-10-23 Toshiki Tsuda , Masaaki Imaizumi

Many theoretical results for the lasso require the samples to be iid. Recent work has provided guarantees for the lasso assuming that the time series is generated by a sparse Vector Auto-Regressive (VAR) model with Gaussian innovations.…

统计理论 · 数学 2019-03-22 Kam Chung Wong , Zifan Li , Ambuj Tewari

In this article we study post-model selection estimators that apply ordinary least squares (OLS) to the model selected by first-step penalized estimators, typically Lasso. It is well known that Lasso can estimate the nonparametric…

统计理论 · 数学 2013-03-21 Alexandre Belloni , Victor Chernozhukov

For consistency (even oracle properties) of estimation and model prediction, almost all existing methods of variable/feature selection critically depend on sparsity of models. However, for ``large $p$ and small $n$" models sparsity…

统计方法学 · 统计学 2010-08-10 Lu Lin , Lixing Zhu , Yujie Gai

We add a set of convex constraints to the lasso to produce sparse interaction models that honor the hierarchy restriction that an interaction only be included in a model if one or both variables are marginally important. We give a precise…

统计方法学 · 统计学 2013-06-20 Jacob Bien , Jonathan Taylor , Robert Tibshirani

Missing values in datasets are common in applied statistics. For regression problems, theoretical work thus far has largely considered the issue of missing covariates as distinct from missing responses. However, in practice, many datasets…

统计理论 · 数学 2026-02-17 Benedict M. Risebrow , Thomas B. Berrett

This note develops an analysis of the Lasso \( \hat b\) in linear models without any sparsity or L1 assumption on the true regression vector, in the proportional regime where dimension \( p \) and sample \( n \) are of the same order. Under…

统计理论 · 数学 2025-01-07 Pierre C. Bellec

Given $n$ noisy samples with $p$ dimensions, where $n \ll p$, we show that the multi-step thresholding procedure based on the Lasso -- we call it the {\it Thresholded Lasso}, can accurately estimate a sparse vector $\beta \in {\mathbb R}^p$…

统计理论 · 数学 2025-10-28 Shuheng Zhou

Graphical models describe associations between variables through the notion of conditional independence. Gaussian graphical models are a widely used class of such models where the relationships are formalized by non-null entries of the…

统计方法学 · 统计学 2023-08-08 Sagnik Bhadury , Riten Mitra , Jeremy T. Gaskins

We consider the use of Bayesian information criteria for selection of the graph underlying an Ising model. In an Ising model, the full conditional distributions of each variable form logistic regression models, and variable selection…

统计理论 · 数学 2015-03-09 Rina Foygel Barber , Mathias Drton

Dantzig selector (DS) and LASSO problems have attracted plenty of attention in statistical learning, sparse data recovery and mathematical optimization. In this paper, we provide a theoretical analysis of the sparse recovery stability of…

统计理论 · 数学 2017-11-13 Yun-Bin Zhao , Duan Li

For multiple index models, it has recently been shown that the sliced inverse regression (SIR) is consistent for estimating the sufficient dimension reduction (SDR) space if and only if $\rho=\lim\frac{p}{n}=0$, where $p$ is the dimension…

统计理论 · 数学 2018-06-19 Qian Lin , Zhigen Zhao , Jun S. Liu

We consider the problem of estimating a low-dimensional parameter in high-dimensional linear regression. Constructing an approximately unbiased estimate of the parameter of interest is a crucial step towards performing statistical…

统计理论 · 数学 2021-07-30 Michael Celentano , Andrea Montanari

Variable (feature, gene, model, which we use interchangeably) selections for regression with high-dimensional BIGDATA have found many applications in bioinformatics, computational biology, image processing, and engineering. One appealing…

机器学习 · 计算机科学 2014-07-29 Zhenqiu Liu , Gang Li

We study full Bayesian procedures for high-dimensional linear regression under sparsity constraints. The prior is a mixture of point masses at zero and continuous distributions. Under compatibility conditions on the design matrix, the…

统计理论 · 数学 2015-10-15 Ismaël Castillo , Johannes Schmidt-Hieber , Aad van der Vaart