An Empirical Study for Vietnamese Constituency Parsing with Pre-training

2020-10-19Unverified0· sign in to hype

Tuan-Vi Tran, Xuan-Thien Pham, Duc-Vu Nguyen, Kiet Van Nguyen, Ngan Luu-Thuy Nguyen

Unverified — Be the first to reproduce this paper.

Abstract

In this work, we use a span-based approach for Vietnamese constituency parsing. Our method follows the self-attention encoder architecture and a chart decoder using a CKY-style inference algorithm. We present analyses of the experiment results of the comparison of our empirical method using pre-training models XLM-Roberta and PhoBERT on both Vietnamese datasets VietTreebank and NIIVTB1. The results show that our model with XLM-Roberta archived the significantly F1-score better than other pre-training models, VietTreebank at 81.19% and NIIVTB1 at 85.70%.

Tasks

Constituency Parsing Decoder Vietnamese Datasets Vietnamese Parsing

An Empirical Study for Vietnamese Constituency Parsing with Pre-training

Abstract

Tasks

Reproductions