English

Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions

Machine Learning 2025-10-28 v2 Machine Learning

Abstract

This paper studies the approximation and generalization abilities of score-based neural network generative models (SGMs) in estimating an unknown distribution P0P_0 from nn i.i.d. observations in dd dimensions. Assuming merely that P0P_0 is α\alpha-sub-Gaussian, we prove that for any time step t[t0,nO(1)]t \in [t_0, n^{\mathcal{O}(1)}], where t0>O(α2n2/dlogn)t_0 > \mathcal{O}(\alpha^2n^{-2/d}\log n), there exists a deep ReLU neural network with width O(n3dlog2n)\leq \mathcal{O}(n^{\frac{3}{d}}\log_2n) and depth O(log2n)\leq \mathcal{O}(\log^2n) that can approximate the scores with O~(n1)\tilde{\mathcal{O}}(n^{-1}) mean square error and achieve a nearly optimal rate of O~(n1t0d/2)\tilde{\mathcal{O}}(n^{-1}t_0^{-d/2}) for score estimation, as measured by the score matching loss. Our framework is universal and can be used to establish convergence rates for SGMs under milder assumptions than previous work. For example, assuming further that the target density function p0p_0 lies in Sobolev or Besov classes, with an appropriately early stopping strategy, we demonstrate that neural network-based SGMs can attain nearly minimax convergence rates up to logarithmic factors. Our analysis removes several crucial assumptions, such as Lipschitz continuity of the score function or a strictly positive lower bound on the target density.

Keywords

Cite

@article{arxiv.2505.10880,
  title  = {Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions},
  author = {Guoji Fu and Wee Sun Lee},
  journal= {arXiv preprint arXiv:2505.10880},
  year   = {2025}
}

Comments

99 pages

R2 v1 2026-06-28T23:35:23.547Z