paper-with-me

홈 › Papers

AGQA 2.0: An Updated Benchmark for Compositional Spatio-Temporal Reasoning

2022-04-12 · Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala

Prior benchmarks have analyzed models' answers to questions about videos in order to measure visual compositional reasoning. Action Genome Question Answering (AGQA) is one such benchmark. AGQA provides a training/test split with balanced answer distributions to reduce the effect of linguistic biases. However, some biases remain in several AGQA categories. We introduce AGQA 2.0, a version of this benchmark with several improvements, most namely a stricter balancing procedure. We then report results on the updated benchmark for all experiments.

📄 PDF Abstract BibTeX arXiv:2204.06105

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

AGQA: A Benchmark for Compositional Spatio-Temporal Reasoning

2021-03-30 · CVPR 2021 1 · Madeleine Grunde-McLaughlin, Ranjay Krishna, Maneesh Agrawala

Visual events are a composition of temporal actions involving actors spatially interacting with objects. When developing computer vision models that can reason about compositional spatio-temporal events, we need benchmar…

Question AnsweringVideo Question AnsweringVisual Reasoning

ANetQA: A Large-scale Benchmark for Fine-grained Compositional Reasoning over Untrimmed Videos

2023-05-04 · CVPR 2023 1 · Zhou Yu, Lixiang Zheng, Zhou Zhao, Fei Wu 외

Building benchmarks to systemically analyze different capabilities of video question answering (VideoQA) models is challenging yet crucial. Existing benchmarks often use non-compositional simple questions and suffer from…

Question AnsweringSpatio-temporal Scene GraphsVideo Question Answering

Neural-Symbolic VideoQA: Learning Compositional Spatio-Temporal Reasoning for Real-world Video Question Answering

2024-04-05 · Lili Liang, Guanglu Sun, Jin Qiu, Lizhong Zhang

Compositional spatio-temporal reasoning poses a significant challenge in the field of video question answering (VideoQA). Existing approaches struggle to establish effective symbolic reasoning structures, which are cruci…

Question AnsweringVideo Question Answering

Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios

2024-11-20 · Shantanu Jaiswal, Debaditya Roy, Basura Fernando, Cheston Tan

Complex visual reasoning and question answering (VQA) is a challenging task that requires compositional multi-step processing and higher-level reasoning capabilities beyond the immediate recognition and localization of o…

Question AnsweringVisual Question Answering (VQA)Visual Reasoning

Measuring Compositional Consistency for Video Question Answering

2022-04-14 · CVPR 2022 1 · Mona Gandhi, Mustafa Omer Gul, Eva Prakash, Madeleine Grunde-McLaughlin 외

Recent video question answering benchmarks indicate that state-of-the-art models struggle to answer compositional questions. However, it remains unclear which types of compositional reasoning cause models to mispredict. …

Question AnsweringVideo Question Answering