paper-with-me

Papers

K-VQG: Knowledge-aware Visual Question Generation for Common-sense Acquisition

2022-03-15 · Kohei Uehara, Tatsuya Harada

Visual Question Generation (VQG) is a task to generate questions from images. When humans ask questions about an image, their goal is often to acquire some new knowledge. However, existing studies on VQG have mainly addressed question generation from answers or question categories, overlooking the objectives of knowledge acquisition. To introduce a knowledge acquisition perspective into VQG, we constructed a novel knowledge-aware VQG dataset called K-VQG. This is the first large, humanly annotated dataset in which questions regarding images are tied to structured knowledge. We also developed a new VQG model that can encode and use knowledge as the target for a question. The experiment results show that our model outperforms existing models on the K-VQG dataset.

📄 PDF Abstract BibTeX arXiv:2203.07890

Code (0)

등록된 구현이 없습니다.

Tasks

Common Sense ReasoningQuestion GenerationQuestion-Generation

Similar Papers 제목 키워드 기반

Questions beyond Pixels: Integrating Commonsense Knowledge in Visual Question Generation for Remote Sensing

2026-02-22 · Siran Li, Li Mi, Javiera Castillo-Navarro, Devis Tuia arxiv

With the rapid development of remote sensing image archives, asking questions about images has become an effective way of gathering specific information or performing semantic image retrieval. However, current automatica…

Question GenerationQuestion AnsweringImage CaptioningImage Retrieval

ConVQG: Contrastive Visual Question Generation with Multimodal Guidance

2024-02-20 · Li Mi, Syrielle Montariol, Javiera Castillo-Navarro, Xianjie Dai 외

Asking questions about visual environments is a crucial way for intelligent agents to understand rich multi-faceted scenes, raising the importance of Visual Question Generation (VQG) systems. Apart from being grounded to…

Question GenerationQuestion-Generation

ConceptBert: Concept-Aware Representation for Visual Question Answering

2020-11-01 · Findings of the Association for Computational Linguistics 2020 · Fran{\c{c}}ois Gard{\`e}res, Maryam Ziaeefard, Baptiste Abeloos, Freddy Lecue

Visual Question Answering (VQA) is a challenging task that has received increasing attention from both the computer vision and the natural language processing communities. A VQA model combines visual and textual features…

Common Sense ReasoningQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Knowledge-aware Visual Question Generation for Remote Sensing Images

2026-02-22 · Siran Li, Li Mi, Javiera Castillo-Navarro, Devis Tuia arxiv

With the rapid development of remote sensing image archives, asking questions about images has become an effective way of gathering specific information or performing image retrieval. However, automatically generated ima…

Question GenerationQuestion AnsweringImage CaptioningImage Retrieval

Visual Commonsense-aware Representation Network for Video Captioning

2022-11-17 · Pengpeng Zeng, Haonan Zhang, Lianli Gao, Xiangpeng Li 외

Generating consecutive descriptions for videos, i.e., Video Captioning, requires taking full advantage of visual representation along with the generation process. Existing video captioning methods focus on making an expl…

Caption GenerationQuestion AnsweringVideo CaptioningVideo Question Answering