paper-with-me

홈 › Papers

An Embarrassingly Simple Approach for Transfer Learning from Pretrained Language Models

2019-02-27 · NAACL 2019 6 · Alexandra Chronopoulou, Christos Baziotis, Alexandros Potamianos

A growing number of state-of-the-art transfer learning methods employ language models pretrained on large generic corpora. In this paper we present a conceptually simple and effective transfer learning approach that addresses the problem of catastrophic forgetting. Specifically, we combine the task-specific optimization function with an auxiliary language model objective, which is adjusted during the training process. This preserves language regularities captured by language models, while enabling sufficient adaptation for solving the target task. Our method does not require pretraining or finetuning separate components of the network and we train our models end-to-end in a single step. We present results on a variety of challenging affective and text classification tasks, surpassing well established transfer learning methods with greater level of complexity.

📄 PDF Abstract BibTeX arXiv:1902.10547

Code (1)

alexandra-chron/siatl 공식 구현 pytorch

Tasks

General ClassificationLanguage ModelingLanguage Modellingtext-classificationText ClassificationTransfer Learning

Similar Papers 제목 키워드 기반

An Embarrassingly Simple Method to Mitigate Undesirable Properties of Pretrained Language Model Tokenizers

2022-05-01 · ACL 2022 5 · Valentin Hofmann, Hinrich Schuetze, Janet Pierrehumbert

We introduce FLOTA (Few Longest Token Approximation), a simple yet effective method to improve the tokenization of pretrained language models (PLMs). FLOTA uses the vocabulary of a standard tokenizer but tries to preserv…

Language ModelingLanguage Modellingtext-classificationText Classification

Embarrassingly Easy Document-Level MT Metrics: How to Convert Any Pretrained Metric Into a Document-Level Metric

2022-09-27 · Giorgos Vernikos, Brian Thompson, Prashant Mathur, Marcello Federico

We hypothesize that existing sentence-level machine translation (MT) metrics become less effective when the human reference contains ambiguities. To verify this hypothesis, we present a very simple method for extending p…

Machine TranslationSentence

Residual Attention: A Simple but Effective Method for Multi-Label Recognition

2021-08-05 · ICCV 2021 10 · Ke Zhu, Jianxin Wu

Multi-label image recognition is a challenging computer vision task of practical use. Progresses in this area, however, are often characterized by complicated methods, heavy computations, and lack of intuitive explanatio…

Multi-Label Image ClassificationMulti-Label Image Recognition

Do It Once: An Embarrassingly Simple Joint Matching Approach to Response Selection

2021-08-01 · Findings (ACL) 2021 8 · Linhao Zhang, Dehong Ma, Sujian Li, Houfeng Wang

Just CHOP: Embarrassingly Simple LLM Compression

2023-05-24 · Ananya Harsh Jha, Tom Sherborne, Evan Pete Walsh, Dirk Groeneveld 외

Large language models (LLMs) enable unparalleled few- and zero-shot reasoning capabilities but at a high computational footprint. A growing assortment of methods for compression promises to reduce the computational burde…

Knowledge DistillationLanguage ModelingLanguage ModellingLarge Language Model+1