Fast mixing for Latent Dirichlet allocation

Johan Jonasson

arXiv:1701.02960·cs.LG·November 2, 2017·1 cites

Fast mixing for Latent Dirichlet allocation

Johan Jonasson

PDF

Open Access

TL;DR

This paper analyzes the mixing time of the Gibbs sampler for a simple case of Latent Dirichlet Allocation, showing it is at most on the order of m^2 log m for two long documents and two words, with implications for more complex cases.

Contribution

It provides the first theoretical bound on the mixing time of the Gibbs sampler for a basic LDA model, demonstrating rapid mixing in a simplified setting.

Findings

01

Mixing time is at most proportional to m^2 log m for the simplified LDA case.

02

The analysis suggests similar mixing behavior in more complex, realistic scenarios.

03

Gibbs sampler converges quickly in the simplified model, indicating potential efficiency in practical applications.

Abstract

Markov chain Monte Carlo (MCMC) algorithms are ubiquitous in probability theory in general and in machine learning in particular. A Markov chain is devised so that its stationary distribution is some probability distribution of interest. Then one samples from the given distribution by running the Markov chain for a "long time" until it appears to be stationary and then collects the sample. However these chains are often very complex and there are no theoretical guarantees that stationarity is actually reached. In this paper we study the Gibbs sampler of the posterior distribution of a very simple case of Latent Dirichlet Allocation, the arguably most well known Bayesian unsupervised learning model for text generation and text classification. It is shown that when the corpus consists of two long documents of equal length $m$ and the vocabulary consists of only two different words, the…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsBayesian Methods and Mixture Models · Data Management and Algorithms · Natural Language Processing Techniques