SOTAVerified

Syntactic Topic Models

2008-12-01NeurIPS 2008Unverified0· sign in to hype

Jordan L. Boyd-Graber, David M. Blei

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

We develop \ (STM), a nonparametric Bayesian model of parsed documents. \ generates words that are both thematically and syntactically constrained, which combines the semantic insights of topic models with the syntactic information available from parse trees. Each word of a sentence is generated by a distribution that combines document-specific topic weights and parse-tree specific syntactic transitions. Words are assumed generated in an order that respects the parse tree. We derive an approximate posterior inference method based on variational methods for hierarchical Dirichlet processes, and we report qualitative and quantitative results on both synthetic data and hand-parsed documents.

Tasks

Reproductions