paper-with-me

홈 › Papers

Trick Me If You Can: Adversarial Writing of Trivia Challenge Questions

2018-07-01 · ACL 2018 7 · Eric Wallace, Jordan Boyd-Graber

Modern question answering systems have been touted as approaching human performance. However, existing question answering datasets are imperfect tests. Questions are written with humans in mind, not computers, and often do not properly expose model limitations. To address this, we develop an adversarial writing setting, where humans interact with trained models and try to break them. This annotation process yields a challenge set, which despite being easy for trivia players to answer, systematically stumps automated question answering systems. Diagnosing model errors on the evaluation data provides actionable insights to explore in developing robust and generalizable question answering systems.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

A novel interface for adversarial trivia question-writing

2024-03-12 · Jason Liu

A critical component when developing question-answering AIs is an adversarial dataset that challenges models to adapt to the complex syntax and reasoning underlying our natural language. Present techniques for procedural…

Question AnsweringSentence

Trick Me If You Can: Human-in-the-loop Generation of Adversarial Examples for Question Answering

2018-09-07 · TACL 2019 3 · Eric Wallace, Pedro Rodriguez, Shi Feng, Ikuya Yamada 외

Adversarial evaluation stress tests a model's understanding of natural language. While past approaches expose superficial patterns, the resulting adversarial examples are limited in complexity and diversity. We propose h…

DiversityInformation RetrievalQuestion AnsweringRetrieval

How the Advent of Ubiquitous Large Language Models both Stymie and Turbocharge Dynamic Adversarial Question Generation

2024-01-20 · Yoo yeon Sung, Ishani Mondal, Jordan Boyd-Graber

Dynamic adversarial question generation, where humans write examples to stump a model, aims to create examples that are realistic and informative. However, the advent of large language models (LLMs) has been a double-edg…

Question GenerationQuestion-GenerationRetrieval

Controllable Decontextualization of Yes/No Question and Answers into Factual Statements

2024-01-18 · Lingbo Mo, Besnik Fetahu, Oleg Rokhlenko, Shervin Malmasi

Yes/No or polar questions represent one of the main linguistic question categories. They consist of a main interrogative clause, for which the answer is binary (assertion or negation). Polar questions and answers (PQA) r…

Negation

TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

2017-05-09 · ACL 2017 7 · Mandar Joshi, Eunsol Choi, Daniel S. Weld, Luke Zettlemoyer

We present TriviaQA, a challenging reading comprehension dataset containing over 650K question-answer-evidence triples. TriviaQA includes 95K question-answer pairs authored by trivia enthusiasts and independently gathere…

Reading ComprehensionSentenceTriviaQA