home › event - improving performance of topic models by variable grouping


Improving performance of topic models by variable grouping
Conferences & Talks



Topic models have a wide range of applications, including modeling of text documents, images, user preferences, product rankings, and many others. However, learning optimal models may be difficult, especially for large problems. The reason is that inference techniques such as Gibbs sampling often converge to suboptimal models due to the abundance of local minima in large datasets. In this paper, we propose a general method of improving the performance of topic models. The method, called 'grouping transform,' works by introducing auxiliary variables which represent assignments of the original model tokens to groups. Using these auxiliary variables, it becomes possible to re-sample an entire group of tokens at a time. This allows the sampler to make larger state space moves. As a result, better models are learned and performance is improved. The proposed ideas are illustrated on several topic models and several text and image datasets. We show that the grouping transform significantly improves performance over standard models.


upcoming events   view all 

The Experience When Business Meets Design
Brian Solis
27 October 2016
PARC Forum  

AI Case Studies: Pushing the Frontiers of Systems Engineering
Tolga Kurtoglu
9 November 2016 | San Francisco, CA
Conferences & Talks  

2016 AIChE Annual Meeting: Energy and Transport Processes
Corie L. Cobb
14 November 2016 - 15 November 2016 | San Francisco, CA
Conferences & Talks  

Printed Electronics USA 2016 - Visit PARC’s Booth #T20
Ross Bringans, Markus Larsson, Janos Veres
16 November 2016 - 17 November 2016 | Santa Clara, CA
Conferences & Talks  

Next Generation Printed Electronics - Invited talk
Janos Veres
28 November 2016 | Boston, MA
Conferences & Talks