Addressing the MFS Bias in WSD systems

2016-05-01LREC 2016Unverified0· sign in to hype

Marten Postma, Ruben Izquierdo, Eneko Agirre, German Rigau, Piek Vossen

Unverified — Be the first to reproduce this paper.

Abstract

Word Sense Disambiguation (WSD) systems tend to have a strong bias towards assigning the Most Frequent Sense (MFS), which results in high performance on the MFS but in a very low performance on the less frequent senses. We addressed the MFS bias in WSD systems by combining the output from a WSD system with a set of mostly static features to create a MFS classifier to decide when to and not to choose the MFS. The output from this MFS classifier, which is based on the Random Forest algorithm, is then used to modify the output from the original WSD system. We applied our classifier to one of the state-of-the-art supervised WSD systems, i.e. IMS, and to of the best state-of-the-art unsupervised WSD systems, i.e. UKB. Our main finding is that we are able to improve the system output in terms of choosing between the MFS and the less frequent senses. When we apply the MFS classifier to fine-grained WSD, we observe an improvement on the less frequent sense cases, whereas we maintain the overall recall.

Tasks

Word Sense Disambiguation

Addressing the MFS Bias in WSD systems

Abstract

Tasks

Reproductions