Human Adversarial QA: Did the Model Understand the Paragraph?
Recently, adversarial attacks have become an important means of gauging the robustness of natural language models as training and testing set methodology has proved inadequate. In this paper we explore an evaluation based on human-in-the-loop adversarial example generation. These adversarial examples aid us in finding the loopholes in the models and give insights into their working. In the published work on adversarial question-answering, perturbations are made on the questions without changing the background context on which the question is based. In the current work, we examine the complementary idea of perturbing the background context while keeping the question constant. We analyze the state-of-the-art language model BERT for the task of question-answering on SQuAD dataset using novel adversarial examples crafted by humans exposing the weaknesses of the model. We present the typology of the successful attacks here as a baseline for stress-testing QA systems.
Code (0)
등록된 구현이 없습니다.
Tasks
Language ModelingLanguage ModellingQuestion AnsweringSimilar Papers 제목 키워드 기반
Adversarial Examples for Evaluating Reading Comprehension Systems
Standard accuracy metrics indicate that reading comprehension systems are making rapid progress, but the extent to which these systems truly understand language remains unclear. To reward systems with real language under…
Question AnsweringReading ComprehensionDROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs
Reading comprehension has recently seen rapid progress, with systems matching humans on the most popular datasets for the task. However, a large body of work has highlighted the brittleness of these systems, showing that…
Question AnsweringReading ComprehensionSemantic ParsingRecurrent Topic-Transition GAN for Visual Paragraph Generation
A natural image usually conveys rich semantic content and can be viewed from different angles. Existing image description methods are largely restricted by small sets of biased visual paragraph annotations, and fail to c…
Generative Adversarial NetworkImage DescriptionImage Paragraph CaptioningSentenceParaCNN: Visual Paragraph Generation via Adversarial Twin Contextual CNNs
Image description generation plays an important role in many real-world applications, such as image retrieval, automatic navigation, and disabled people support. A well-developed task of image description generation is i…
Image CaptioningImage DescriptionImage RetrievalRetrieval+1Power in Numbers: Robust reading comprehension by finetuning with four adversarial sentences per example
Recent models have achieved human level performance on the Stanford Question Answering Dataset when using F1 scores to evaluate the reading comprehension task. Yet, teaching machines to comprehend text has not been solve…
Question AnsweringReading ComprehensionSentence