paper-with-me

홈 › Papers

XLM-E: Cross-lingual Language Model Pre-training via ELECTRA

2021-11-16 · ACL ARR November 2021 11 · Anonymous

In this paper, we introduce ELECTRA-style tasks to cross-lingual language model pre-training. Specifically, we present two pre-training tasks, namely multilingual replaced token detection, and translation replaced token detection. Besides, we pretrain the model, named as XLM-E, on both multilingual and parallel corpora. Our model outperforms the baseline models on various cross-lingual understanding tasks with much less computation cost. Moreover, analysis shows that XLM-E tends to obtain better cross-lingual transferability.

📄 PDF Abstract BibTeX

Code (1)

pwc-1/Paper-9/tree/main/1/xlm mindspore

Tasks

Language ModelingLanguage ModellingTranslation

Similar Papers 제목 키워드 기반

XLM-E: Cross-lingual Language Model Pre-training via ELECTRA

2021-06-30 · ACL 2022 5 · Zewen Chi, Shaohan Huang, Li Dong, Shuming Ma 외

In this paper, we introduce ELECTRA-style tasks to cross-lingual language model pre-training. Specifically, we present two pre-training tasks, namely multilingual replaced token detection, and translation replaced token …

Language ModelingLanguage ModellingTranslationZero-Shot Cross-Lingual Transfer

DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

2021-11-18 · Pengcheng He, Jianfeng Gao, Weizhu Chen

This paper presents a new pre-trained language model, DeBERTaV3, which improves the original DeBERTa model by replacing mask language modeling (MLM) with replaced token detection (RTD), a more sample-efficient pre-traini…

Language ModelingLanguage ModellingNatural Language InferenceNatural Language Understanding+2

MCL@IITK at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation using Augmented Data, Signals, and Transformers

2021-04-04 · SEMEVAL 2021 · Rohan Gupta, Jay Mundra, Deepak Mahajan, Ashutosh Modi

In this work, we present our approach for solving the SemEval 2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation (MCL-WiC). The task is a sentence pair classification problem where the goal is to …

Cross-Lingual TransferSentenceSentence-Pair ClassificationTask 2+1

Pre-training and Evaluating Transformer-based Language Models for Icelandic

2022-06-01 · LREC 2022 6 · Jón Guðnason, Hrafn Loftsson

In this paper, we evaluate several Transformer-based language models for Icelandic on four downstream tasks: Part-of-Speech tagging, Named Entity Recognition. Dependency Parsing, and Automatic Text Summarization. We pre-…

Dependency Parsingnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

BioELECTRA:Pretrained Biomedical text Encoder using Discriminators

2021-06-11 · ACL Anthology 2021 6 · Kamal raj Kanakarajan, Bhuvana Kundumani, Malaikannan Sankarasubbu

Recent advancements in pretraining strategies in NLP have shown a significant improvement in the performance of models on various text mining tasks. We apply ‘replaced token detection’ pretraining technique proposed by E…

ArticlesLanguage ModelingLanguage ModellingMedical Named Entity Recognition+3