The Importance of Being Recurrent for Modeling Hierarchical Structure

2018-03-09EMNLP 2018Code Available0· sign in to hype

Ke Tran, Arianna Bisazza, Christof Monz

Code Available — Be the first to reproduce this paper.

Code

github.com/ketranm/fan_vs_rnn
OfficialIn paperpytorch★ 0

Abstract

Recent work has shown that recurrent neural networks (RNNs) can implicitly capture and exploit hierarchical information when trained to solve common natural language processing tasks such as language modeling (Linzen et al., 2016) and neural machine translation (Shi et al., 2016). In contrast, the ability to model structured data with non-recurrent neural networks has received little attention despite their success in many NLP tasks (Gehring et al., 2017; Vaswani et al., 2017). In this work, we compare the two architectures---recurrent versus non-recurrent---with respect to their ability to model hierarchical structure and find that recurrency is indeed important for this purpose.

Tasks

Language Modeling Language Modelling Machine Translation Translation

The Importance of Being Recurrent for Modeling Hierarchical Structure

Code

Abstract

Tasks

Reproductions