Understanding Priors in Bayesian Neural Networks at the Unit Level
Abstract
We investigate deep Bayesian neural networks with Gaussian weight priors and a class of ReLU-like nonlinearities. Bayesian neural networks with Gaussian priors are well known to induce an L2, "weight decay", regularization. Our results characterize a more intricate regularization effect at the level of the unit activations. Our main result establishes that the induced prior distribution on the units before and after activation becomes increasingly heavy-tailed with the depth of the layer. We show that first layer units are Gaussian, second layer units are sub-exponential, and units in deeper layers are characterized by sub-Weibull distributions. Our results provide new theoretical insight on deep Bayesian neural networks, which we corroborate with simulation experiments.
Keywords
Cite
@article{arxiv.1810.05193,
title = {Understanding Priors in Bayesian Neural Networks at the Unit Level},
author = {Mariia Vladimirova and Jakob Verbeek and Pablo Mesejo and Julyan Arbel},
journal= {arXiv preprint arXiv:1810.05193},
year = {2019}
}
Comments
10 pages, 5 figures, ICML'19 conference