paper-with-me

MSVD-QA

홈페이지 · 논문 61편

The MSVD-QA dataset is a Video Question Answering (VideoQA) dataset. It is based on the existing Microsoft Research Video Description (MSVD) dataset, which consists of about 120K sentences describing more than 2,000 video snippets. In the MSVD-QA dataset, Question-Answer (QA) pairs are generated from these descriptions. The dataset is mainly used in video captioning experiments but due to its large data size, it is also used for VideoQA. It contains 1970 video clips and approximately 50.5K QA pairs.

벤치마크

Zero-Shot Video Question Answer on MSVD-QA 결과 56개
Visual Question Answering (VQA) on MSVD-QA 결과 36개
Visual Question Answering on MSVD-QA 결과 4개
Video Question Answering on MSVD-QA 결과 1개
Zero-Shot Learning on MSVD-QA 결과 1개