SOTAVerified

NLPOP: a Dataset for Popularity Prediction of Promoted NLP Research on Twitter

2022-05-01WASSA (ACL) 2022Code Available0· sign in to hype

Leo Obadić, Martin Tutek, Jan Šnajder

Code Available — Be the first to reproduce this paper.

Reproduce

Code

Abstract

Twitter has slowly but surely established itself as a forum for disseminating, analysing and promoting NLP research. The trend of researchers promoting work not yet peer-reviewed (preprints) by posting concise summaries presented itself as an opportunity to collect and combine multiple modalities of data. In scope of this paper, we (1) construct a dataset of Twitter threads in which researchers promote NLP preprints and (2) evaluate whether it is possible to predict the popularity of a thread based on the content of the Twitter thread, paper content and user metadata. We experimentally show that it is possible to predict popularity of threads promoting research based on their content, and that predictive performance depends on modelling textual input, indicating that the dataset could present value for related areas of NLP research such as citation recommendation and abstractive summarization.

Tasks

Reproductions