paper-with-me

홈 › Papers

Super-CLEVR: A Virtual Benchmark to Diagnose Domain Robustness in Visual Reasoning

2022-12-01 · CVPR 2023 1 · Zhuowan Li, Xingrui Wang, Elias Stengel-Eskin, Adam Kortylewski, Wufei Ma, Benjamin Van Durme, Alan Yuille

Visual Question Answering (VQA) models often perform poorly on out-of-distribution data and struggle on domain generalization. Due to the multi-modal nature of this task, multiple factors of variation are intertwined, making generalization difficult to analyze. This motivates us to introduce a virtual benchmark, Super-CLEVR, where different factors in VQA domain shifts can be isolated in order that their effects can be studied independently. Four factors are considered: visual complexity, question redundancy, concept distribution and concept compositionality. With controllably generated data, Super-CLEVR enables us to test VQA methods in situations where the test data differs from the training data along each of these axes. We study four existing methods, including two neural symbolic methods NSCL and NSVQA, and two non-symbolic methods FiLM and mDETR; and our proposed method, probabilistic NSVQA (P-NSVQA), which extends NSVQA with uncertainty reasoning. P-NSVQA outperforms other methods on three of the four domain shift factors. Our results suggest that disentangling reasoning and perception, combined with probabilistic uncertainty, form a strong VQA model that is more robust to domain shifts. The dataset and code are released at https://github.com/Lizw14/Super-CLEVR.

📄 PDF Abstract BibTeX arXiv:2212.00259

Code (2)

lizw14/super-clevr 공식 구현
XingruiWang/superclevr-3D-question

Tasks

Domain GeneralizationQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)Visual Reasoning

Similar Papers 제목 키워드 기반

ClevrTex: A Texture-Rich Benchmark for Unsupervised Multi-Object Segmentation

2021-11-19 · Laurynas Karazija, Iro Laina, Christian Rupprecht

There has been a recent surge in methods that aim to decompose and segment scenes into multiple objects in an unsupervised manner, i.e., unsupervised multi-object segmentation. Performing such a task is a long-standing g…

SegmentationSemantic SegmentationUnsupervised Object Segmentation

No Representation Rules Them All in Category Discovery

2023-11-28 · NeurIPS 2023 11

In this paper we tackle the problem of Generalized Category Discovery (GCD). Specifically, given a dataset with labelled and unlabelled images, the task is to cluster all images in the unlabelled subset, whether or not t…

AllClusteringRepresentation Learning

Distilling Answer Set Programming Theories from Large Language Models

2026-07-30 · Nelson Higuera Ruiz, Markus Hofmarcher, Claudiu Leoveanu-Condrei arxiv

Writing Answer Set Programming (ASP) theories from scratch is a difficult and time-consuming task. We take a neurosymbolic approach to study whether a model can distill complete and correct theories, given a fixed agent …

Measuring CLEVRness: Blackbox testing of Visual Reasoning Models

2022-02-24 · Spyridon Mouselinos, Henryk Michalewski, Mateusz Malinowski

How can we measure the reasoning capabilities of intelligence systems? Visual question answering provides a convenient framework for testing the model's abilities by interrogating the model through questions about the sc…

BenchmarkingDiagnosticQuestion AnsweringVisual Question Answering+2

Measuring CLEVRness: Black-box Testing of Visual Reasoning Models

2021-09-29 · ICLR 2022 4 · Spyridon Mouselinos, Henryk Michalewski, Mateusz Malinowski

How to measure the reasoning capabilities of intelligence systems? Visual question answering provides a convenient framework for testing the model's abilities by interrogating the model through questions about the scene.…

BenchmarkingDiagnosticQuestion AnsweringVisual Question Answering+2