Distant Supervision from Disparate Sources for Low-Resource Part-of-Speech Tagging
2018-08-29EMNLP 2018Code Available0· sign in to hype
Barbara Plank, Željko Agić
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/bplank/bilstm-auxOfficialIn papernone★ 0
Abstract
We introduce DsDs: a cross-lingual neural part-of-speech tagger that learns from disparate sources of distant supervision, and realistically scales to hundreds of low-resource languages. The model exploits annotation projection, instance selection, tag dictionaries, morphological lexicons, and distributed representations, all in a uniform framework. The approach is simple, yet surprisingly effective, resulting in a new state of the art without access to any gold annotated data.