Designing Templates for Eliciting Commonsense Knowledge from Pretrained Sequence-to-Sequence Models

2020-12-01COLING 2020Unverified0· sign in to hype

Jheng-Hong Yang, Sheng-Chieh Lin, Rodrigo Nogueira, Ming-Feng Tsai, Chuan-Ju Wang, Jimmy Lin

Unverified — Be the first to reproduce this paper.

Abstract

While internalized ``implicit knowledge'' in pretrained transformers has led to fruitful progress in many natural language understanding tasks, how to most effectively elicit such knowledge remains an open question. Based on the text-to-text transfer transformer (T5) model, this work explores a template-based approach to extract implicit knowledge for commonsense reasoning on multiple-choice (MC) question answering tasks. Experiments on three representative MC datasets show the surprisingly good performance of our simple template, coupled with a logit normalization technique for disambiguation. Furthermore, we verify that our proposed template can be easily extended to other MC tasks with contexts such as supporting facts in open-book question answering settings. Starting from the MC task, this work initiates further research to find generic natural language templates that can effectively leverage stored knowledge in pretrained models.

Tasks

Multiple-choice Natural Language Understanding Open-Ended Question Answering Question Answering

Designing Templates for Eliciting Commonsense Knowledge from Pretrained Sequence-to-Sequence Models

Abstract

Tasks

Reproductions