English

Understanding Priors in Bayesian Neural Networks at the Unit Level

Machine Learning 2019-05-13 v2 Machine Learning

Abstract

We investigate deep Bayesian neural networks with Gaussian weight priors and a class of ReLU-like nonlinearities. Bayesian neural networks with Gaussian priors are well known to induce an L2, "weight decay", regularization. Our results characterize a more intricate regularization effect at the level of the unit activations. Our main result establishes that the induced prior distribution on the units before and after activation becomes increasingly heavy-tailed with the depth of the layer. We show that first layer units are Gaussian, second layer units are sub-exponential, and units in deeper layers are characterized by sub-Weibull distributions. Our results provide new theoretical insight on deep Bayesian neural networks, which we corroborate with simulation experiments.

Keywords

Cite

@article{arxiv.1810.05193,
  title  = {Understanding Priors in Bayesian Neural Networks at the Unit Level},
  author = {Mariia Vladimirova and Jakob Verbeek and Pablo Mesejo and Julyan Arbel},
  journal= {arXiv preprint arXiv:1810.05193},
  year   = {2019}
}

Comments

10 pages, 5 figures, ICML'19 conference

R2 v1 2026-06-23T04:36:51.586Z