SOTAVerified

Appendix - Recommended Statistical Significance Tests for NLP Tasks

2018-09-05Code Available0· sign in to hype

Rotem Dror, Roi Reichart

Code Available — Be the first to reproduce this paper.

Reproduce

Code

Abstract

Statistical significance testing plays an important role when drawing conclusions from experimental results in NLP papers. Particularly, it is a valuable tool when one would like to establish the superiority of one algorithm over another. This appendix complements the guide for testing statistical significance in NLP presented in dror2018hitchhiker by proposing valid statistical tests for the common tasks and evaluation measures in the field.

Tasks

Reproductions