English

High-dimensional online learning via asynchronous decomposition: Non-divergent results, dynamic regularization, and beyond

Machine Learning 2026-03-24 v1 Machine Learning

Abstract

Existing high-dimensional online learning methods often face the challenge that their error bounds, or per-batch sample sizes, diverge as the number of data batches increases. To address this issue, we propose an asynchronous decomposition framework that leverages summary statistics to construct a surrogate score function for current-batch learning. This framework is implemented via a dynamic-regularized iterative hard thresholding algorithm, providing a computationally and memory-efficient solution for sparse online optimization. We provide a unified theoretical analysis that accounts for both the streaming computational error and statistical accuracy, establishing that our estimator maintains non-divergent error bounds and 0\ell_0 sparsity across all batches. Furthermore, the proposed estimator adaptively achieves additional gains as batches accumulate, attaining the oracle accuracy as if the entire historical dataset were accessible and the true support were known. These theoretical properties are further illustrated through an example of the generalized linear model.

Keywords

Cite

@article{arxiv.2603.20696,
  title  = {High-dimensional online learning via asynchronous decomposition: Non-divergent results, dynamic regularization, and beyond},
  author = {Shixiang Liu and Zhifan Li and Hanming Yang and Jianxin Yin},
  journal= {arXiv preprint arXiv:2603.20696},
  year   = {2026}
}

Comments

41 pages, 1 figure

R2 v1 2026-07-01T11:31:09.120Z