K-VQG: Knowledge-aware Visual Question Generation for Common-sense Acquisition
Visual Question Generation (VQG) is a task to generate questions from images. When humans ask questions about an image, their goal is often to acquire some new knowledge. However, existing studies on VQG have mainly addressed question generation from answers or question categories, overlooking the objectives of knowledge acquisition. To introduce a knowledge acquisition perspective into VQG, we constructed a novel knowledge-aware VQG dataset called K-VQG. This is the first large, humanly annotated dataset in which questions regarding images are tied to structured knowledge. We also developed a new VQG model that can encode and use knowledge as the target for a question. The experiment results show that our model outperforms existing models on the K-VQG dataset.
Code (0)
등록된 구현이 없습니다.
Tasks
Common Sense ReasoningQuestion GenerationQuestion-GenerationSimilar Papers 제목 키워드 기반
Questions beyond Pixels: Integrating Commonsense Knowledge in Visual Question Generation for Remote Sensing
With the rapid development of remote sensing image archives, asking questions about images has become an effective way of gathering specific information or performing semantic image retrieval. However, current automatica…
Question GenerationQuestion AnsweringImage CaptioningImage RetrievalConVQG: Contrastive Visual Question Generation with Multimodal Guidance
Asking questions about visual environments is a crucial way for intelligent agents to understand rich multi-faceted scenes, raising the importance of Visual Question Generation (VQG) systems. Apart from being grounded to…
Question GenerationQuestion-GenerationConceptBert: Concept-Aware Representation for Visual Question Answering
Visual Question Answering (VQA) is a challenging task that has received increasing attention from both the computer vision and the natural language processing communities. A VQA model combines visual and textual features…
Common Sense ReasoningQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)Knowledge-aware Visual Question Generation for Remote Sensing Images
With the rapid development of remote sensing image archives, asking questions about images has become an effective way of gathering specific information or performing image retrieval. However, automatically generated ima…
Question GenerationQuestion AnsweringImage CaptioningImage RetrievalVisual Commonsense-aware Representation Network for Video Captioning
Generating consecutive descriptions for videos, i.e., Video Captioning, requires taking full advantage of visual representation along with the generation process. Existing video captioning methods focus on making an expl…
Caption GenerationQuestion AnsweringVideo CaptioningVideo Question Answering