English

Cryptographic Hardness of Learning Halfspaces with Massart Noise

Machine Learning 2022-07-29 v1 Computational Complexity Data Structures and Algorithms

Abstract

We study the complexity of PAC learning halfspaces in the presence of Massart noise. In this problem, we are given i.i.d. labeled examples (x,y)RN×{±1}(\mathbf{x}, y) \in \mathbb{R}^N \times \{ \pm 1\}, where the distribution of x\mathbf{x} is arbitrary and the label yy is a Massart corruption of f(x)f(\mathbf{x}), for an unknown halfspace f:RN{±1}f: \mathbb{R}^N \to \{ \pm 1\}, with flipping probability η(x)η<1/2\eta(\mathbf{x}) \leq \eta < 1/2. The goal of the learner is to compute a hypothesis with small 0-1 error. Our main result is the first computational hardness result for this learning problem. Specifically, assuming the (widely believed) subexponential-time hardness of the Learning with Errors (LWE) problem, we show that no polynomial-time Massart halfspace learner can achieve error better than Ω(η)\Omega(\eta), even if the optimal 0-1 error is small, namely OPT=2logc(N)\mathrm{OPT} = 2^{-\log^{c} (N)} for any universal constant c(0,1)c \in (0, 1). Prior work had provided qualitatively similar evidence of hardness in the Statistical Query model. Our computational hardness result essentially resolves the polynomial PAC learnability of Massart halfspaces, by showing that known efficient learning algorithms for the problem are nearly best possible.

Keywords

Cite

@article{arxiv.2207.14266,
  title  = {Cryptographic Hardness of Learning Halfspaces with Massart Noise},
  author = {Ilias Diakonikolas and Daniel M. Kane and Pasin Manurangsi and Lisheng Ren},
  journal= {arXiv preprint arXiv:2207.14266},
  year   = {2022}
}