SOTAVerified

AGQA 2.0: An Updated Benchmark for Compositional Spatio-Temporal Reasoning

2022-04-12Unverified0· sign in to hype

Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

Prior benchmarks have analyzed models' answers to questions about videos in order to measure visual compositional reasoning. Action Genome Question Answering (AGQA) is one such benchmark. AGQA provides a training/test split with balanced answer distributions to reduce the effect of linguistic biases. However, some biases remain in several AGQA categories. We introduce AGQA 2.0, a version of this benchmark with several improvements, most namely a stricter balancing procedure. We then report results on the updated benchmark for all experiments.

Tasks

Reproductions