arXiv · 2204.06105
AGQA 2.0: An Updated Benchmark for Compositional Spatio-Temporal Reasoning
Abstract
Prior benchmarks have analyzed models' answers to questions about videos in order to measure visual compositional reasoning. Action Genome Question Answering (AGQA) is one such benchmark. AGQA provides a training/test split with balanced answer distributions to reduce the effect of linguistic biases. However, some biases remain in several AGQA categories. We introduce AGQA 2.0, a version of this benchmark with several improvements, most namely a stricter balancing procedure. We then report results on the updated benchmark for all experiments.
Explore related subjects
Keep this discovery
Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala. 2022-04-12. AGQA 2.0: An Updated Benchmark for Compositional Spatio-Temporal Reasoning. https://arxiv.org/abs/2204.06105
Cite the original work for its findings. Save a collection to share your selection of sources.