Using Paraphrasing and Memory-Augmented Models to Combat Data Sparsity in Question Interpretation with a Virtual Patient Dialogue System

2018-06-01WS 2018Unverified0· sign in to hype

Lifeng Jin, David King, Amad Hussein, Michael White, Douglas Danforth

Unverified — Be the first to reproduce this paper.

Abstract

When interpreting questions in a virtual patient dialogue system one must inevitably tackle the challenge of a long tail of relatively infrequently asked questions. To make progress on this challenge, we investigate the use of paraphrasing for data augmentation and neural memory-based classification, finding that the two methods work best in combination. In particular, we find that the neural memory-based approach not only outperforms a straight CNN classifier on low frequency questions, but also takes better advantage of the augmented data created by paraphrasing, together yielding a nearly 10\% absolute improvement in accuracy on the least frequently asked questions.

Tasks

Data Augmentation General Classification One-Shot Learning

Using Paraphrasing and Memory-Augmented Models to Combat Data Sparsity in Question Interpretation with a Virtual Patient Dialogue System

Abstract

Tasks

Reproductions