paper-with-me

홈 › Papers

mLUKE: The Power of Entity Representations in Multilingual Pretrained Language Models

2021-10-15 · ACL 2022 5 · Ryokan Ri, Ikuya Yamada, Yoshimasa Tsuruoka

Recent studies have shown that multilingual pretrained language models can be effectively improved with cross-lingual alignment information from Wikipedia entities. However, existing methods only exploit entity information in pretraining and do not explicitly use entities in downstream tasks. In this study, we explore the effectiveness of leveraging entity representations for downstream cross-lingual tasks. We train a multilingual language model with 24 languages with entity representations and show the model consistently outperforms word-based pretrained models in various cross-lingual transfer tasks. We also analyze the model and the key insight is that incorporating entity representations into the input allows us to extract more language-agnostic features. We also evaluate the model with a multilingual cloze prompt task with the mLAMA dataset. We show that entity-based prompt elicits correct factual knowledge more likely than using only word representations. Our source code and pretrained models are available at https://github.com/studio-ousia/luke.

📄 PDF Abstract BibTeX arXiv:2110.08151

Code (4)

studio-ousia/luke 공식 구현 pytorch
2023-MindSpore-1/ms-code-5/tree/main/luke mindspore
pwc-1/Paper-5/tree/main/mluke mindspore
pwc-1/Paper-9/tree/main/2/mluke mindspore

Tasks

Cross-Lingual Question AnsweringCross-Lingual TransferLanguage ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

ClassBases at CASE-2022 Multilingual Protest Event Detection Tasks: Multilingual Protest News Detection and Automatically Replicating Manually Created Event Datasets

2023-01-16 · Peratham Wiriyathammabhum

In this report, we describe our ClassBases submissions to a shared task on multilingual protest event detection. For the multilingual protest news detection, we participated in subtask-1, subtask-2, and subtask-4, which …

ClassificationDocument ClassificationEvent DetectionSentence+3

Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection

2024-12-29 · Kalin Kopanov

The integration of advanced Natural Language Processing (NLP) methodologies and Large Language Models (LLMs) has significantly enhanced the extraction and analysis of geospatial data from multilingual texts, impacting se…

Sequence Tagging with Contextual and Non-Contextual Subword Representations: A Multilingual Evaluation

2019-06-04 · ACL 2019 7 · Benjamin Heinzerling, Michael Strube

Pretrained contextual and non-contextual subword embeddings have become available in over 250 languages, allowing massively multilingual NLP. However, while there is no dearth of pretrained embeddings, the distinct lack …

Multilingual Named Entity RecognitionMultilingual NLPnamed-entity-recognitionNamed Entity Recognition+2

Human-Annotated NER Dataset for the Kyrgyz Language

2025-09-23 · Timur Turatali, Anton Alekseev, Gulira Jumalieva, Gulnara Kabaeva 외 arxiv

We introduce KyrgyzNER, the first manually annotated named entity recognition dataset for the Kyrgyz language. Comprising 1,499 news articles from the 24.KG news portal, the dataset contains 10,900 sentences and 39,075 e…

Aligning Multilingual Embeddings for Improved Code-switched Natural Language Understanding

2022-10-01 · COLING 2022 10 · Barah Fazili, Preethi Jyothi

Multilingual pretrained models, while effective on monolingual data, need additional training to work well with code-switched text. In this work, we present a novel idea of training multilingual models with alignment obj…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Natural Language Understanding+2