paper-with-me

홈 › Papers

Learning from Task Descriptions

2020-11-16 · EMNLP 2020 11 · Orion Weller, Nicholas Lourie, Matt Gardner, Matthew E. Peters

Typically, machine learning systems solve new tasks by training on thousands of examples. In contrast, humans can solve new tasks by reading some instructions, with perhaps an example or two. To take a step toward closing this gap, we introduce a framework for developing NLP systems that solve new tasks after reading their descriptions, synthesizing prior work in this area. We instantiate this framework with a new English language dataset, ZEST, structured for task-oriented evaluation on unseen tasks. Formulating task descriptions as questions, we ensure each is general enough to apply to many possible inputs, thus comprehensively evaluating a model's ability to solve each task. Moreover, the dataset's structure tests specific types of systematic generalization. We find that the state-of-the-art T5 model achieves a score of 12% on ZEST, leaving a significant challenge for NLP researchers.

📄 PDF Abstract BibTeX arXiv:2011.08115

Code (1)

allenai/zest tf

Tasks

Systematic Generalization

Methods 이 논문이 사용한 방법론

Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Inverse Square Root Schedule Inverse Square Root is a learning rate schedule 1 / $\sqrt{\max\left(n, k\right)}$ where $n$ is the current training iteration and $k$ is the number of warm-up steps. This…
SentencePiece 설명 없음
Attention 설명 없음
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Adafactor Adafactor is a stochastic optimization method based on Adam that reduces memory usage while retaining the empirical benefits of…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

Varying image description tasks: spoken versus written descriptions

2018-08-01 · COLING 2018 8 · Emiel van Miltenburg, Ruud Koolen, Emiel Krahmer

Automatic image description systems are commonly trained and evaluated on written image descriptions. At the same time, these systems are often used to provide spoken descriptions (e.g. for visually impaired users) throu…

Image Description

CityFlow-NL: Tracking and Retrieval of Vehicles at City Scale by Natural Language Descriptions

2021-01-12 · Qi Feng, Vitaly Ablavsky, Stan Sclaroff

Natural Language (NL) descriptions can be one of the most convenient or the only way to interact with systems built to understand and detect city scale traffic patterns and vehicle-related events. In this paper, we exten…

Multi-Object TrackingObject TrackingRetrievalTemporal Localization

When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions

2025-07-27 · Maya Larbi, Amal Akli, Mike Papadakis, Rihab Bouyousfi 외 arxiv

Large Language Models (LLMs) have demonstrated impressive performance in code generation tasks under idealized conditions, where task descriptions are clear and precise. However, in practice, task descriptions frequently…

Code Generation

MathAlign: Linking Formula Identifiers to their Contextual Natural Language Descriptions

2020-05-01 · LREC 2020 5 · Maria Alexeeva, Rebecca Sharp, Marco A. Valenzuela-Esc{\'a}rcega, Jennifer Kadowaki 외

Extending machine reading approaches to extract mathematical concepts and their descriptions is useful for a variety of tasks, ranging from mathematical information retrieval to increasing accessibility of scientific doc…

Information RetrievalReading ComprehensionRetrieval

ACCoRD: A Multi-Document Approach to Generating Diverse Descriptions of Scientific Concepts

2022-05-14 · Sonia K. Murthy, Kyle Lo, Daniel King, Chandra Bhagavatula 외

Systems that can automatically define unfamiliar terms hold the promise of improving the accessibility of scientific texts, especially for readers who may lack prerequisite background knowledge. However, current systems …