FVQA 2.0: Introducing Adversarial Samples into Fact-based Visual Question Answering
The widely used Fact-based Visual Question Answering (FVQA) dataset contains visually-grounded questions that require information retrieval using common sense knowledge graphs to answer. It has been observed that the original dataset is highly imbalanced and concentrated on a small portion of its associated knowledge graph. We introduce FVQA 2.0 which contains adversarial variants of test questions to address this imbalance. We show that systems trained with the original FVQA train sets can be vulnerable to adversarial samples and we demonstrate an augmentation scheme to reduce this vulnerability without human annotations.
Code (0)
등록된 구현이 없습니다.
Tasks
Common Sense ReasoningInformation RetrievalKnowledge GraphsQuestion AnsweringRetrievalVisual Question AnsweringVisual Question Answering (VQA)Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
FVQA: Fact-based Visual Question Answering
Visual Question Answering (VQA) has attracted a lot of attention in both Computer Vision and Natural Language Processing communities, not least because it offers insight into the relationships between two important sourc…
Common Sense ReasoningQuestion AnsweringTripletVisual Question Answering+1Seeing is Knowing! Fact-based Visual Question Answering using Knowledge Graph Embeddings
Fact-based Visual Question Answering (FVQA), a challenging variant of VQA, requires a QA-system to include facts from a diverse knowledge graph (KG) in its reasoning process to produce an answer. Large KGs, especially co…
Common Sense ReasoningKnowledge Graph EmbeddingsQuestion AnsweringRetrieval+3FVQ: A Large-Scale Dataset and A LMM-based Method for Face Video Quality Assessment
Face video quality assessment (FVQA) deserves to be explored in addition to general video quality assessment (VQA), as face videos are the primary content on social media platforms and human visual system (HVS) is partic…
Video Quality AssessmentVisual Question Answering (VQA)High-Fidelity Video Quality Assessment with VQA-Specific Saliency
No-reference video quality assessment (NR VQA) has recently seen promising progress with deep learning. However, video data is inherently large, and processing them with deep models incurs high computational cost. This c…
Video Quality AssessmentMucko: Multi-Layer Cross-Modal Knowledge Reasoning for Fact-based Visual Question Answering
Fact-based Visual Question Answering (FVQA) requires external knowledge beyond visible content to answer questions about an image, which is challenging but indispensable to achieve general VQA. One limitation of existing…
Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)