paper-with-me

SQuAD

Stanford Question Answering Dataset

홈페이지 · 논문 151편

The Stanford Question Answering Dataset (SQuAD) is a collection of question-answer pairs derived from Wikipedia articles. In SQuAD, the correct answers of questions can be any sequence of tokens in the given text. Because the questions and answers are produced by humans through crowdsourcing, it is more diverse than some other question-answering datasets. SQuAD 1.1 contains 107,785 question-answer pairs on 536 articles. SQuAD2.0 (open-domain SQuAD, SQuAD-Open), the latest version, combines the 100,000 questions in SQuAD1.1 with over 50,000 un-answerable questions written adversarially by crowdworkers in forms that are similar to the answerable ones. Source: Deep Learning Based Text Classification: A Comprehensive Review Image Source: https://rajpurkar.github.io/SQuAD-explorer/explore/v2.0/dev/Prime_number.html

Texts English

벤치마크

Question Answering on SQuAD2.0 결과 286개
Question Answering on SQuAD1.1 결과 213개
Question Answering on SQuAD1.1 dev 결과 55개
Question Answering on SQuAD2.0 dev 결과 13개
Question Generation on SQuAD1.1 결과 13개
Data-free Knowledge Distillation on SQuAD 결과 8개
Open-Domain Question Answering on SQuAD1.1 dev 결과 8개
Open-Domain Question Answering on SQuAD1.1 결과 6개
Question Answering on SQuAD 결과 2개
Question Generation on SQuAD 결과 2개