Statistical Foundation of Variational Bayes Neural Networks

Shrijita Bhattacharya; Tapabrata Maiti

arXiv:2006.15786·stat.ML·June 30, 2020

Statistical Foundation of Variational Bayes Neural Networks

Shrijita Bhattacharya, Tapabrata Maiti

PDF

TL;DR

This paper establishes the theoretical validity of variational Bayes methods for neural networks by proving posterior consistency and analyzing convergence rates, thus enhancing their reliability in complex data scenarios.

Contribution

It provides the first theoretical proof of posterior consistency for variational Bayes in neural networks, outlining conditions for convergence and guiding prior construction.

Findings

01

Proves posterior consistency of variational Bayes in neural networks.

02

Analyzes the influence of scale parameters on convergence rates.

03

Provides guidelines for prior distribution design in Bayesian neural networks.

Abstract

Despite the popularism of Bayesian neural networks in recent years, its use is somewhat limited in complex and big data situations due to the computational cost associated with full posterior evaluations. Variational Bayes (VB) provides a useful alternative to circumvent the computational cost and time complexity associated with the generation of samples from the true posterior using Markov Chain Monte Carlo (MCMC) techniques. The efficacy of the VB methods is well established in machine learning literature. However, its potential broader impact is hindered due to a lack of theoretical validity from a statistical perspective. However there are few results which revolve around the theoretical properties of VB, especially in non-parametric problems. In this paper, we establish the fundamental result of posterior consistency for the mean-field variational posterior (VP) for a feed-forward…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.