paper-with-me

Papers

Cross-Task Generalization via Natural Language Crowdsourcing Instructions

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Humans (e.g., crowdworkers) have a remarkable ability in solving different tasks, by simply reading textual instructions that define them and looking at a few examples. Despite the success of the conventional supervised learning on individual datasets, such models often struggle with generalization across tasks (e.g., a question-answering system cannot solve classification tasks). A long-standing challenge in AI is to build a model that learns a new task by understanding the human-readable instructions that define it. To study this, we introduce NATURAL INSTRUCTIONS, a dataset of 61 distinct tasks, their human-authored instructions, and 193k task instances (input-output pairs). The instructions are obtained from crowdsourcing instructions used to create existing NLP datasets and mapped to a unified schema. Using this meta-dataset, we measure cross-task generalization by training models on seen tasks and measuring generalization to the remaining unseen ones. We adopt generative pre-trained language models to encode task-specific instructions along with input and generate task output. Our results indicate that models benefit from instructions when evaluated in terms of generalization to unseen tasks (19% better for models utilizing instructions). These models, however, are far behind an estimated performance upperbound indicating significant room for more progress in this direction.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Cross-Task Generalization via Natural Language Crowdsourcing Instructions

2021-04-18 · ACL 2022 5 · Swaroop Mishra, Daniel Khashabi, Chitta Baral, Hannaneh Hajishirzi

Humans (e.g., crowdworkers) have a remarkable ability in solving different tasks, by simply reading textual instructions that define them and looking at a few examples. Despite the success of the conventional supervised …

Question Answering

Learning Lexical Entries for Robotic Commands using Crowdsourcing

2016-09-08 · Junjie Hu, Jean Oh, Anatole Gershman

Robotic commands in natural language usually contain various spatial descriptions that are semantically similar but syntactically different. Mapping such syntactic variants into semantic concepts that can be understood b…

Machine TranslationTranslation

Crowdsourcing Natural Language Data at Scale: A Hands-On Tutorial

2021-06-01 · NAACL 2021 4 · Alexey Drutsa, Dmitry Ustalov, Valentina Fedorova, Olga Megorskaya 외

In this tutorial, we present a portion of unique industry experience in efficient natural language data annotation via crowdsourcing shared by both leading researchers and engineers from Yandex. We will make an introduct…

Design Choices for Crowdsourcing Implicit Discourse Relations: Revealing the Biases Introduced by Task Design

2023-04-03 · Valentina Pyatkin, Frances Yung, Merel C. J. Scholman, Reut Tsarfaty 외

Disagreement in natural language annotation has mostly been studied from a perspective of biases introduced by the annotators and the annotation frameworks. Here, we propose to analyze another source of bias: task design…

Crowdsourcing Lexical Diversity

2024-10-30 · Hadi Khalilia, Jahna Otterbacher, Gabor Bella, Rusma Noortyani 외

Lexical-semantic resources (LSRs), such as online lexicons or wordnets, are fundamental for natural language processing applications. In many languages, however, such resources suffer from quality issues: incorrect entri…

Diversity