AMR Similarity Metrics from Principles

2020-01-29Code Available1· sign in to hype

Juri Opitz, Letitia Parcalabescu, Anette Frank

Code Available — Be the first to reproduce this paper.

Code

github.com/Heidelberg-NLP/amr-metric-suite
OfficialIn papernone★ 11
github.com/flipz357/amr-metric-suite
none★ 18
github.com/fliegenpilz357/amr-metric-suite
none★ 18

Abstract

Different metrics have been proposed to compare Abstract Meaning Representation (AMR) graphs. The canonical Smatch metric (Cai and Knight, 2013) aligns the variables of two graphs and assesses triple matches. The recent SemBleu metric (Song and Gildea, 2019) is based on the machine-translation metric Bleu (Papineni et al., 2002) and increases computational efficiency by ablating the variable-alignment. In this paper, i) we establish criteria that enable researchers to perform a principled assessment of metrics comparing meaning representations like AMR; ii) we undertake a thorough analysis of Smatch and SemBleu where we show that the latter exhibits some undesirable properties. For example, it does not conform to the identity of indiscernibles rule and introduces biases that are hard to control; iii) we propose a novel metric S^2match that is more benevolent to only very slight meaning deviations and targets the fulfilment of all established criteria. We assess its suitability and show its advantages over Smatch and SemBleu.

Tasks

Abstract Meaning Representation Computational Efficiency Graph Matching Machine Translation Translation

Benchmark Results

Dataset	Model	Metric	Claimed	Verified	Status
RARE	S2match	Spearman Correlation	94.11	—	Unverified

AMR Similarity Metrics from Principles

Code

Abstract

Tasks

Benchmark Results

Reproductions