Undirected Graphical Models as Approximate Posteriors

Arash Vahdat; Evgeny Andriyash; William G. Macready

arXiv:1901.03440·stat.ML·June 9, 2020·5 cites

Undirected Graphical Models as Approximate Posteriors

Arash Vahdat, Evgeny Andriyash, William G. Macready

PDF

Open Access 3 Repos 1 Video

TL;DR

This paper introduces a method to improve variational autoencoders by using undirected graphical models as approximate posteriors, leveraging MCMC-based gradient estimation for training, resulting in better performance than directed models.

Contribution

It develops an efficient training approach for undirected approximate posteriors in VAEs using backpropagation through MCMC updates, enhancing generative quality.

Findings

01

Undirected models outperform directed graphical models in VAEs.

02

The proposed gradient estimator enables effective training of discrete VAEs.

03

Implementation demonstrates improved generative performance.

Abstract

The representation of the approximate posterior is a critical aspect of effective variational autoencoders (VAEs). Poor choices for the approximate posterior have a detrimental impact on the generative performance of VAEs due to the mismatch with the true posterior. We extend the class of posterior models that may be learned by using undirected graphical models. We develop an efficient method to train undirected approximate posteriors by showing that the gradient of the training objective with respect to the parameters of the undirected posterior can be computed by backpropagation through Markov chain Monte Carlo updates. We apply these gradient estimators for training discrete VAEs with Boltzmann machines as approximate posteriors and demonstrate that undirected models outperform previous results obtained using directed graphical models. Our implementation is available at…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

Undirected Graphical Models as Approximate Posteriors· slideslive

Taxonomy

TopicsGenerative Adversarial Networks and Image Synthesis · Model Reduction and Neural Networks · Gaussian Processes and Bayesian Inference