English

SQ Lower Bounds for Learning Mixtures of Linear Classifiers

Machine Learning 2023-10-19 v1 Data Structures and Algorithms Statistics Theory Machine Learning Statistics Theory

Abstract

We study the problem of learning mixtures of linear classifiers under Gaussian covariates. Given sample access to a mixture of rr distributions on Rn\mathbb{R}^n of the form (x,y)(\mathbf{x},y_{\ell}), [r]\ell\in [r], where xN(0,In)\mathbf{x}\sim\mathcal{N}(0,\mathbf{I}_n) and y=sign(v,x)y_\ell=\mathrm{sign}(\langle\mathbf{v}_\ell,\mathbf{x}\rangle) for an unknown unit vector v\mathbf{v}_\ell, the goal is to learn the underlying distribution in total variation distance. Our main result is a Statistical Query (SQ) lower bound suggesting that known algorithms for this problem are essentially best possible, even for the special case of uniform mixtures. In particular, we show that the complexity of any SQ algorithm for the problem is npoly(1/Δ)log(r)n^{\mathrm{poly}(1/\Delta) \log(r)}, where Δ\Delta is a lower bound on the pairwise 2\ell_2-separation between the v\mathbf{v}_\ell's. The key technical ingredient underlying our result is a new construction of spherical designs that may be of independent interest.

Keywords

Cite

@article{arxiv.2310.11876,
  title  = {SQ Lower Bounds for Learning Mixtures of Linear Classifiers},
  author = {Ilias Diakonikolas and Daniel M. Kane and Yuxin Sun},
  journal= {arXiv preprint arXiv:2310.11876},
  year   = {2023}
}

Comments

To appear in NeurIPS 2023

R2 v1 2026-06-28T12:54:15.499Z