paper-with-me

홈 › Papers

mOthello: When Do Cross-Lingual Representation Alignment and Cross-Lingual Transfer Emerge in Multilingual Models?

2024-04-18 · Tianze Hua, Tian Yun, Ellie Pavlick

Many pretrained multilingual models exhibit cross-lingual transfer ability, which is often attributed to a learned language-neutral representation during pretraining. However, it remains unclear what factors contribute to the learning of a language-neutral representation, and whether the learned language-neutral representation suffices to facilitate cross-lingual transfer. We propose a synthetic task, Multilingual Othello (mOthello), as a testbed to delve into these two questions. We find that: (1) models trained with naive multilingual pretraining fail to learn a language-neutral representation across all input languages; (2) the introduction of "anchor tokens" (i.e., lexical items that are identical across languages) helps cross-lingual representation alignment; and (3) the learning of a language-neutral representation alone is not sufficient to facilitate cross-lingual transfer. Based on our findings, we propose a novel approach - multilingual pretraining with unified output space - that both induces the learning of language-neutral representation and facilitates cross-lingual transfer.

📄 PDF Abstract BibTeX arXiv:2404.12444

Code (1)

ethahtz/multilingual_othello 공식 구현 pytorch

Tasks

Cross-Lingual Transfer

Similar Papers 제목 키워드 기반

Exploring the Relationship between Alignment and Cross-lingual Transfer in Multilingual Transformers

2023-06-05 · Félix Gaschi, Patricio Cerda, Parisa Rastin, Yannick Toussaint

Without any explicit cross-lingual training data, multilingual language models can achieve cross-lingual transfer. One common way to improve this transfer is to perform realignment steps before fine-tuning, i.e., to trai…

Cross-Lingual TransferPOSPOS TaggingXLM-R

AlignX: Advancing Multilingual Large Language Models with Multilingual Representation Alignment

2025-09-29 · Mengyu Bu, Shaolei Zhang, Zhongjun He, Hua Wu 외 arxiv

Multilingual large language models (LLMs) possess impressive multilingual understanding and generation capabilities. However, their performance and cross-lingual alignment often lag for non-dominant languages. A common s…

High-Dimensional Interlingual Representations of Large Language Models

2025-03-14 · Bryan Wilie, Samuel Cahyawijaya, Junxian He, Pascale Fung

Large language models (LLMs) trained on massive multilingual datasets hint at the formation of interlingual constructs--a shared subspace in the representation space. However, evidence regarding this phenomenon is mixed,…

Limitations and Challenges of Unsupervised Cross-lingual Pre-training

2022-09-01 · AMTA 2022 9 · Martín Quesada Zaragoza, Francisco Casacuberta

Cross-lingual alignment methods for monolingual language representations have received notable attention in recent years. However, their use in machine translation pre-training remains scarce. This work tries to shed lig…

Machine TranslationTranslation

Cross-Lingual Representation Alignment Through Contrastive Image-Caption Tuning

2025-05-19 · Nathaniel Krasner, Nicholas Lanuzo, Antonios Anastasopoulos

Multilingual alignment of sentence representations has mostly required bitexts to bridge the gap between languages. We investigate whether visual information can bridge this gap instead. Image caption datasets are very e…

Natural Language UnderstandingRetrievalSentence