English

Nash: Neural Adaptive Shrinkage for Structured High-Dimensional Regression

Machine Learning 2026-05-19 v2 Machine Learning

Abstract

Sparse linear regression is a fundamental tool in data analysis. However, traditional approaches often fall short when covariates exhibit structure or arise from heterogeneous sources. In biomedical applications, covariates may stem from distinct modalities or be structured according to an underlying graph. We introduce \textit{Neural Adaptive Shrinkage} (Nash), a unified framework that integrates covariate-specific side information into sparse regression via neural networks. Nash adaptively modulates penalties on a per-covariate basis, learning to tailor regularization without cross-validation. We use a \textit{split variational empirical Bayes} algorithm that decouples prior learning from posterior inference, reducing the M-step from O(p)\mathcal{O}(p) neural-network passes per sweep to a single batched pass, a \textit{74 to 106x wall-clock speedup} over previously proposed coordinate ascent CAVI for p between 10210^2 and 10410^4. Experiments on real data demonstrate that Nash improves accuracy and adaptability over existing methods.

Keywords

Cite

@article{arxiv.2505.11143,
  title  = {Nash: Neural Adaptive Shrinkage for Structured High-Dimensional Regression},
  author = {William R. P. Denault},
  journal= {arXiv preprint arXiv:2505.11143},
  year   = {2026}
}
R2 v1 2026-06-28T23:35:51.508Z