arXiv · 1810.05193
Understanding Priors in Bayesian Neural Networks at the Unit Level
Abstract
We investigate deep Bayesian neural networks with Gaussian weight priors and a class of ReLU-like nonlinearities. Bayesian neural networks with Gaussian priors are well known to induce an L2, "weight decay", regularization. Our results characterize a more intricate regularization effect at the level of the unit activations. Our main result establishes that the induced prior distribution on the units before and after activation becomes increasingly heavy-tailed with the depth of the layer. We show that first layer units are Gaussian, second layer units are sub-exponential, and units in deeper layers are characterized by sub-Weibull distributions. Our results provide new theoretical insight on deep Bayesian neural networks, which we corroborate with simulation experiments.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Mariia Vladimirova, Jakob Verbeek, Pablo Mesejo, Julyan Arbel. 2019-05-10. Understanding Priors in Bayesian Neural Networks at the Unit Level. https://arxiv.org/abs/1810.05193
Cite the original work for its findings. Save a collection to share your selection of sources.