paper-with-me

홈 › Papers

What Gives the Answer Away? Question Answering Bias Analysis on Video QA Datasets

2020-07-07 · Jianing Yang, Yuying Zhu, Yongxin Wang, Ruitao Yi, Amir Zadeh, Louis-Philippe Morency

Question answering biases in video QA datasets can mislead multimodal model to overfit to QA artifacts and jeopardize the model's ability to generalize. Understanding how strong these QA biases are and where they come from helps the community measure progress more accurately and provide researchers insights to debug their models. In this paper, we analyze QA biases in popular video question answering datasets and discover pretrained language models can answer 37-48% questions correctly without using any multimodal context information, far exceeding the 20% random guess baseline for 5-choose-1 multiple-choice questions. Our ablation study shows biases can come from annotators and type of questions. Specifically, annotators that have been seen during training are better predicted by the model and reasoning, abstract questions incur more biases than factual, direct questions. We also show empirically that using annotator-non-overlapping train-test splits can reduce QA biases for video QA datasets.

📄 PDF Abstract BibTeX arXiv:2007.03626

Code (0)

등록된 구현이 없습니다.

Tasks

Multiple-choiceQuestion AnsweringVideo Question Answering

Similar Papers 제목 키워드 기반

Analysis of Wikipedia-based Corpora for Question Answering

2018-01-06 · Tomasz Jurczyk, Amit Deshmane, Jinho D. Choi

This paper gives comprehensive analyses of corpora based on Wikipedia for several tasks in question answering. Four recent corpora are collected,WikiQA, SelQA, SQuAD, and InfoQA, and first analyzed intrinsically by conte…

Question AnsweringRetrieval

Inverse Visual Question Answering: A New Benchmark and VQA Diagnosis Tool

2018-03-16 · Feng Liu, Tao Xiang, Timothy M. Hospedales, Wankou Yang 외

In recent years, visual question answering (VQA) has become topical. The premise of VQA's significance as a benchmark in AI, is that both the image and textual question need to be well understood and mutually grounded in…

Question AnsweringReinforcement LearningVisual Question AnsweringVisual Question Answering (VQA)

Where To Look: Focus Regions for Visual Question Answering

2015-11-23 · CVPR 2016 6 · Kevin J. Shih, Saurabh Singh, Derek Hoiem

We present a method that learns to answer visual questions by selecting image regions relevant to the text-based query. Our method exhibits significant improvements in answering questions such as "what color," where it i…

Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)

SPARQL query generation for complex question answering with BERT and BiLSTM-based model

2020-06-17 · Proceedings of the International Conference “Dialogue 2020” 2020 6 · Evseev D. A., Arkhipov M. Yu.

In this paper we describe question answering system for answering of complex questions over Wikidata knowledge base. Unlike simple questions, which require extraction of single fact from the knowledge base, complex quest…

Knowledge Base Question AnsweringQuestion AnsweringSemantic ParsingTriplet

The Impact of Answers in Referential Visual Dialog

2021-10-01 · ReInAct 2021 10 · Mauricio Mazuecos, Patrick Blackburn, Luciana Benotti

In the visual dialog task GuessWhat?! two players maintain a dialog in order to identify a secret object in an image. Computationally, this is modeled using a question generation module and a guesser module for the quest…

Question GenerationQuestion-GenerationVisual Dialog