SOTAVerified

Active DOP: A constituency treebank annotation tool with online learning

2018-08-01COLING 2018Code Available0· sign in to hype

Andreas van Cranenburgh

Code Available — Be the first to reproduce this paper.

Reproduce

Code

Abstract

We present a language-independent treebank annotation tool supporting rich annotations with discontinuous constituents and function tags. Candidate analyses are generated by an exemplar-based parsing model that immediately learns from each new annotated sentence during annotation. This makes it suitable for situations in which only a limited seed treebank is available, or a radically different domain is being annotated. The tool offers the possibility to experiment with and evaluate active learning methods to speed up annotation in a naturalistic setting, i.e., measuring actual annotation costs and tracking specific user interactions. The code is made available under the GNU GPL license at https://github.com/andreasvc/activedop.

Tasks

Reproductions