paper-with-me

홈 › Papers

NSP-BERT: A Prompt-based Few-Shot Learner through an Original Pre-training Task —— Next Sentence Prediction

2022-10-01 · COLING 2022 10 · Yi Sun, Yu Zheng, Chao Hao, Hangping Qiu

Using prompts to utilize language models to perform various downstream tasks, also known as prompt-based learning or prompt-learning, has lately gained significant success in comparison to the pre-train and fine-tune paradigm. Nonetheless, virtually most prompt-based methods are token-level such as PET based on mask language model (MLM). In this paper, we attempt to accomplish several NLP tasks in the zero-shot and few-shot scenarios using a BERT original pre-training task abandoned by RoBERTa and other models——Next Sentence Prediction (NSP). Unlike token-level techniques, our sentence-level prompt-based method NSP-BERT does not need to fix the length of the prompt or the position to be predicted, allowing it to handle tasks such as entity linking with ease. NSP-BERT can be applied to a variety of tasks based on its properties. We present an NSP-tuning approach with binary cross-entropy loss for single-sentence classification tasks that is competitive compared to PET and EFL. By continuing to train BERT on RoBERTa’s corpus, the model’s performance improved significantly, which indicates that the pre-training corpus is another important determinant of few-shot besides model size and prompt method.

📄 PDF Abstract BibTeX

Code (1)

sunyilgdx/prompts4keras 공식 구현 tf

Tasks

Entity LinkingLanguage ModelingLanguage ModellingPrompt LearningSentenceSentence Classification

Similar Papers 제목 키워드 기반

NSP-BERT: A Prompt-based Few-Shot Learner Through an Original Pre-training Task--Next Sentence Prediction

2021-09-08 · Yi Sun, Yu Zheng, Chao Hao, Hangping Qiu

Using prompts to utilize language models to perform various downstream tasks, also known as prompt-based learning or prompt-learning, has lately gained significant success in comparison to the pre-train and fine-tune par…

Entity LinkingLanguage ModelingLanguage ModellingPrompt Learning+3

NSP-BERT: A Prompt-based Zero-Shot Learner Through an Original Pre-training Task —— Next Sentence Prediction

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Using prompts to utilize language models to perform various downstream tasks, also known as prompt-based learning or prompt-learning, has lately gained significant success in comparison to the pre-train and fine-tune par…

Entity LinkingLanguage ModelingLanguage ModellingPrompt Learning+2

Tutorials on Stance Detection using Pre-trained Language Models: Fine-tuning BERT and Prompting Large Language Models

2023-07-28 · Yun-Shiuan Chuang

This paper presents two self-contained tutorials on stance detection in Twitter data using BERT fine-tuning and prompting large language models (LLMs). The first tutorial explains BERT architecture and tokenization, guid…

Stance Detection

ELECTRA is a Zero-Shot Learner, Too

2022-07-17 · Shiwen Ni, Hung-Yu Kao

Recently, for few-shot or even zero-shot learning, the new paradigm "pre-train, prompt, and predict" has achieved remarkable achievements compared with the "pre-train, fine-tune" paradigm. After the success of prompt-bas…

Language ModelingLanguage ModellingPrompt LearningSST-2+1

Making Small Language Models Better Few-Shot Learners

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Large-scale language models coupled with prompts have shown remarkable performance on few-shot learning. However, through systematic experiments, we find that the few-shot performance of small language models is poor, an…

Few-Shot LearningKnowledge DistillationSentenceSentence Classification