English

Revisiting mean estimation over $\ell_p$ balls: Is the MLE optimal?

Statistics Theory 2025-07-02 v2 Information Theory math.IT Statistics Theory

Abstract

We revisit the problem of mean estimation in the Gaussian sequence model with p\ell_p constraints for p[0,]p \in [0, \infty]. We demonstrate two phenomena for the behavior of the maximum likelihood estimator (MLE), which depend on the noise level, the radius of the (quasi)norm constraint, the dimension, and the norm index pp. First, if pp lies between 00 and 1+Θ(1logd)1 + \Theta(\tfrac{1}{\log d}), inclusive, or if it is greater than or equal to 22, the MLE is minimax rate-optimal for all noise levels and all constraint radii. On the other hand, for the remaining norm indices -- namely, if pp lies between 1+Θ(1logd)1 + \Theta(\tfrac{1}{\log d}) and 22 -- here is a more striking behavior: the MLE is minimax rate-suboptimal, despite its nonlinearity in the observations, for essentially all noise levels and constraint radii for which nonlinear estimates are necessary for minimax-optimal estimation. Our results imply that when given nn independent and identically distributed Gaussian samples, the MLE can be suboptimal by a polynomial factor in the sample size. Our lower bounds are constructive: whenever the MLE is rate-suboptimal, we provide explicit instances on which the MLE provably incurs suboptimal risk. Finally, in the non-convex case -- namely when p<1p < 1 -- we develop sharp local Gaussian width bounds, which may be of independent interest.

Keywords

Cite

@article{arxiv.2506.10354,
  title  = {Revisiting mean estimation over $\ell_p$ balls: Is the MLE optimal?},
  author = {Liviu Aolaritei and Michael I. Jordan and Reese Pathak and Annie Ulichney},
  journal= {arXiv preprint arXiv:2506.10354},
  year   = {2025}
}

Comments

43 pages, 3 figures