Multi-class Hierarchical Question Classification for Multiple Choice Science Exams

2019-08-15LREC 2020Code Available0· sign in to hype

Dongfang Xu, Peter Jansen, Jaycie Martin, Zhengnan Xie, Vikas Yadav, Harish Tayyar Madabushi, Oyvind Tafjord, Peter Clark

arXiv PDF

Code Available — Be the first to reproduce this paper.

Reproduce

Code

github.com/cognitiveailab/questionclassification
tf★ 0

Abstract

Prior work has demonstrated that question classification (QC), recognizing the problem domain of a question, can help answer it more accurately. However, developing strong QC algorithms has been hindered by the limited size and complexity of annotated data available. To address this, we present the largest challenge dataset for QC, containing 7,787 science exam questions paired with detailed classification labels from a fine-grained hierarchical taxonomy of 406 problem domains. We then show that a BERT-based model trained on this dataset achieves a large (+0.12 MAP) gain compared with previous methods, while also achieving state-of-the-art performance on benchmark open-domain and biomedical QC datasets. Finally, we show that using this model's predictions of question topic significantly improves the accuracy of a question answering system by +1.7% P@1, with substantial future gains possible as QC performance improves.

Tasks

Classification General Classification Multiple-choice Question Answering

Multi-class Hierarchical Question Classification for Multiple Choice Science Exams

Code

Abstract

Tasks

Reproductions