BasqueGLUE: A Natural Language Understanding Benchmark for Basque

2022-06-01LREC 2022Code Available0· sign in to hype

Gorka Urbizu, Iñaki San Vicente, Xabier Saralegi, Rodrigo Agerri, Aitor Soroa

Code Available — Be the first to reproduce this paper.

Code

github.com/elhuyar/basqueglue
OfficialIn papernone★ 1

Abstract

Natural Language Understanding (NLU) technology has improved significantly over the last few years and multitask benchmarks such as GLUE are key to evaluate this improvement in a robust and general way. These benchmarks take into account a wide and diverse set of NLU tasks that require some form of language understanding, beyond the detection of superficial, textual clues. However, they are costly to develop and language-dependent, and therefore they are only available for a small number of languages. In this paper, we present BasqueGLUE, the first NLU benchmark for Basque, a less-resourced language, which has been elaborated from previously existing datasets and following similar criteria to those used for the construction of GLUE and SuperGLUE. We also report the evaluation of two state-of-the-art language models for Basque on BasqueGLUE, thus providing a strong baseline to compare upon. BasqueGLUE is freely available under an open license.

Tasks

Natural Language Understanding

BasqueGLUE: A Natural Language Understanding Benchmark for Basque

Code

Abstract

Tasks

Reproductions