Predicting and Using Target Length in Neural Machine Translation

2020-12-01Asian Chapter of the Association for Computational LinguisticsUnverified0· sign in to hype

Zijian Yang, Yingbo Gao, Weiyue Wang, Hermann Ney

Unverified — Be the first to reproduce this paper.

Abstract

Attention-based encoder-decoder models have achieved great success in neural machine translation tasks. However, the lengths of the target sequences are not explicitly predicted in these models. This work proposes length prediction as an auxiliary task and set up a sub-network to obtain the length information from the encoder. Experimental results show that the length prediction sub-network brings improvements over the strong baseline system and that the predicted length can be used as an alternative to length normalization during decoding.

Tasks

Decoder Machine Translation Prediction Translation

Predicting and Using Target Length in Neural Machine Translation

Abstract

Tasks

Reproductions