Modeling Sense Structure in Word Usage Graphs with the Weighted Stochastic Block Model

2021-08-01Joint Conference on Lexical and Computational SemanticsCode Available0· sign in to hype

Dominik Schlechtweg, Enrique Castaneda, Jonas Kuhn, Sabine Schulte im Walde

Code Available — Be the first to reproduce this paper.

Code

github.com/kicasta/modeling_wugs_wsbm
OfficialIn papernone★ 1

Abstract

We suggest to model human-annotated Word Usage Graphs capturing fine-grained semantic proximity distinctions between word uses with a Bayesian formulation of the Weighted Stochastic Block Model, a generative model for random graphs popular in biology, physics and social sciences. By providing a probabilistic model of graded word meaning we aim to approach the slippery and yet widely used notion of word sense in a novel way. The proposed framework enables us to rigorously compare models of word senses with respect to their fit to the data. We perform extensive experiments and select the empirically most adequate model.

Tasks

Stochastic Block Model

Modeling Sense Structure in Word Usage Graphs with the Weighted Stochastic Block Model

Code

Abstract

Tasks

Reproductions