SOTAVerified

TLT-CRF: A Lexicon-supported Morphological Tagger for Latin Based on Conditional Random Fields

2016-05-01LREC 2016Unverified0· sign in to hype

Tim vor der Br{\"u}ck, Alex Mehler, er

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

We present a morphological tagger for Latin, called TTLab Latin Tagger based on Conditional Random Fields (TLT-CRF) which uses a large Latin lexicon. Beyond Part of Speech (PoS), TLT-CRF tags eight inflectional categories of verbs, adjectives or nouns. It utilizes a statistical model based on CRFs together with a rule interpreter that addresses scenarios of sparse training data. We present results of evaluating TLT-CRF to answer the question what can be learnt following the paradigm of 1st order CRFs in conjunction with a large lexical resource and a rule interpreter. Furthermore, we investigate the contigency of representational features and targeted parts of speech to learn about selective features.

Tasks

Reproductions