Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Lee Gao

Communication-Free Parallel Supervised Topic Models

Aug 10, 2017

Lee Gao, Ronghuo Zheng

Figure 1 for Communication-Free Parallel Supervised Topic Models

Figure 2 for Communication-Free Parallel Supervised Topic Models

Figure 3 for Communication-Free Parallel Supervised Topic Models

Figure 4 for Communication-Free Parallel Supervised Topic Models

Abstract:Embarrassingly (communication-free) parallel Markov chain Monte Carlo (MCMC) methods are commonly used in learning graphical models. However, MCMC cannot be directly applied in learning topic models because of the quasi-ergodicity problem caused by multimodal distribution of topics. In this paper, we develop an embarrassingly parallel MCMC algorithm for sLDA. Our algorithm works by switching the order of sampled topics combination and labeling variable prediction in sLDA, in which it overcomes the quasi-ergodicity problem because high-dimension topics that follow a multimodal distribution are projected into one-dimension document labels that follow a unimodal distribution. Our empirical experiments confirm that the out-of-sample prediction performance using our embarrassingly parallel algorithm is comparable to non-parallel sLDA while the computation time is significantly reduced.

* 8 pages, 7 figures

Via

Access Paper or Ask Questions