SOTAVerified

Learning Language from a Large (Unannotated) Corpus

2014-01-14Code Available0· sign in to hype

Linas Vepstas, Ben Goertzel

Code Available — Be the first to reproduce this paper.

Reproduce

Code

Abstract

A novel approach to the fully automated, unsupervised extraction of dependency grammars and associated syntax-to-semantic-relationship mappings from large text corpora is described. The suggested approach builds on the authors' prior work with the Link Grammar, RelEx and OpenCog systems, as well as on a number of prior papers and approaches from the statistical language learning literature. If successful, this approach would enable the mining of all the information needed to power a natural language comprehension and generation system, directly from a large, unannotated corpus.

Reproductions