paper-with-me

홈 › Papers

Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor

2022-12-19 · Or Honovich, Thomas Scialom, Omer Levy, Timo Schick

Instruction tuning enables pretrained language models to perform new tasks from inference-time natural language descriptions. These approaches rely on vast amounts of human supervision in the form of crowdsourced datasets or user interactions. In this work, we introduce Unnatural Instructions: a large dataset of creative and diverse instructions, collected with virtually no human labor. We collect 64,000 examples by prompting a language model with three seed examples of instructions and eliciting a fourth. This set is then expanded by prompting the model to rephrase each instruction, creating a total of approximately 240,000 examples of instructions, inputs, and outputs. Experiments show that despite containing a fair amount of noise, training on Unnatural Instructions rivals the effectiveness of training on open-source manually-curated datasets, surpassing the performance of models such as T0++ and Tk-Instruct across various benchmarks. These results demonstrate the potential of model-generated data as a cost-effective alternative to crowdsourcing for dataset expansion and diversification.

📄 PDF Abstract BibTeX arXiv:2212.09689

Code (3)

orhonovich/unnatural-instructions 공식 구현
baai-zlab/coig
flagopen/flaginstruct

Tasks

Language ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Self-Instruct: Aligning Language Models with Self-Generated Instructions

2022-12-20 · Yizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu 외

Large "instruction-tuned" language models (i.e., finetuned to respond to instructions) have demonstrated a remarkable ability to generalize zero-shot to new tasks. Nevertheless, they depend heavily on human-written instr…

Instruction FollowingLanguage Modelling

Unnatural Error Correction: GPT-4 Can Almost Perfectly Handle Unnatural Scrambled Text

2023-11-30 · Qi Cao, Takeshi Kojima, Yutaka Matsuo, Yusuke Iwasawa

While Large Language Models (LLMs) have achieved remarkable performance in many tasks, much about their inner workings remains unclear. In this study, we present novel experimental insights into the resilience of LLMs, p…

Balancing Shared Autonomy with Human-Robot Communication

2018-05-20 · Rosario Scalise, Yonatan Bisk, Maxwell Forbes, Daqing Yi 외

Robotic agents that share autonomy with a human should leverage human domain knowledge and account for their preferences when completing a task. This extra knowledge can dramatically improve plan efficiency and user-sati…

Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning

2024-02-09 · Shivalika Singh, Freddie Vargus, Daniel Dsouza, Börje F. Karlsson 외

Datasets are foundational to many breakthroughs in modern artificial intelligence. Many recent achievements in the space of natural language processing (NLP) can be attributed to the finetuning of pre-trained models on a…

Instruction FollowingLanguage ModelingLanguage ModellingLarge Language Model

Privacy-Preserving Instructions for Aligning Large Language Models

2024-02-21 · Da Yu, Peter Kairouz, Sewoong Oh, Zheng Xu

Service providers of large language model (LLM) applications collect user instructions in the wild and use them in further aligning LLMs with users' intentions. These instructions, which potentially contain sensitive inf…

Language ModelingLanguage ModellingLarge Language ModelPrivacy Preserving