English

Computation of the Maximum Likelihood estimator in low-rank Factor Analysis

Optimization and Control 2018-01-19 v1 Computation Machine Learning

Abstract

Factor analysis, a classical multivariate statistical technique is popularly used as a fundamental tool for dimensionality reduction in statistics, econometrics and data science. Estimation is often carried out via the Maximum Likelihood (ML) principle, which seeks to maximize the likelihood under the assumption that the positive definite covariance matrix can be decomposed as the sum of a low rank positive semidefinite matrix and a diagonal matrix with nonnegative entries. This leads to a challenging rank constrained nonconvex optimization problem. We reformulate the low rank ML Factor Analysis problem as a nonlinear nonsmooth semidefinite optimization problem, study various structural properties of this reformulation and propose fast and scalable algorithms based on difference of convex (DC) optimization. Our approach has computational guarantees, gracefully scales to large problems, is applicable to situations where the sample covariance matrix is rank deficient and adapts to variants of the ML problem with additional constraints on the problem parameters. Our numerical experiments demonstrate the significant usefulness of our approach over existing state-of-the-art approaches.

Keywords

Cite

@article{arxiv.1801.05935,
  title  = {Computation of the Maximum Likelihood estimator in low-rank Factor Analysis},
  author = {Koulik Khamaru and Rahul Mazumder},
  journal= {arXiv preprint arXiv:1801.05935},
  year   = {2018}
}

Comments

22 pages, 4 figures

R2 v1 2026-06-22T23:48:29.615Z