paper-with-me

홈 › Papers

A Three-Pronged Approach to Cross-Lingual Adaptation with Multilingual LLMs

2024-06-25 · Vaibhav Singh, Amrith Krishna, Karthika NJ, Ganesh Ramakrishnan

Low-resource languages, by its very definition, tend to be under represented in the pre-training corpora of Large Language Models. In this work, we investigate three low-resource cross-lingual approaches that enable an LLM adapt to tasks in previously unseen languages. Llama-2 is an LLM where Indic languages, among many other language families, contribute to less than $0.005\%$ of the total $2$ trillion token pre-training corpora. In this work, we experiment with the English-dominated Llama-2 for cross-lingual transfer to three Indic languages, Bengali, Hindi, and Tamil as target languages. We study three approaches for cross-lingual transfer, under ICL and fine-tuning. One, we find that adding additional supervisory signals via a dominant language in the LLM, leads to improvements, both under in-context learning and fine-tuning. Two, adapting the target languages to word reordering may be beneficial under ICL, but its impact diminishes with fine tuning. Finally, continued pre-training in one low-resource language can improve model performance for other related low-resource languages.

📄 PDF Abstract BibTeX arXiv:2406.17377

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual TransferIn-Context Learning

Similar Papers 제목 키워드 기반

Accounting for Language Effect in the Evaluation of Cross-lingual AMR Parsers

2022-10-01 · COLING 2022 10 · Shira Wein, Nathan Schneider

Cross-lingual Abstract Meaning Representation (AMR) parsers are currently evaluated in comparison to gold English AMRs, despite parsing a language other than English, due to the lack of multilingual AMR evaluation metric…

Abstract Meaning RepresentationSentence

Task-Conditional Flow Matching for Balanced Multilingual Text Embedding Adaptation

2026-08-06 · Tirth Bhatt, Naren Kumar S, Mayank Singh arxiv

Multilingual text embedding models are commonly adapted using a single training objective across diverse tasks, despite different tasks requiring fundamentally different optimization strategies. We introduce Task-Conditi…

Everything Is All It Takes: A Multipronged Strategy for Zero-Shot Cross-Lingual Information Extraction

2021-09-14 · EMNLP 2021 11 · Mahsa Yarmohammadi, Shijie Wu, Marc Marone, Haoran Xu 외

Zero-shot cross-lingual information extraction (IE) describes the construction of an IE model for some target language, given existing annotations exclusively in some other language, typically English. While the advance …

AllDependency ParsingEvent Extractionnamed-entity-recognition+3

Multilingual Pre-training with Language and Task Adaptation for Multilingual Text Style Transfer

2022-03-16 · ACL 2022 5 · Huiyuan Lai, Antonio Toral, Malvina Nissim

We exploit the pre-trained seq2seq model mBART for multilingual text style transfer. Using machine translated data as well as gold aligned English sentences yields state-of-the-art results in the three target languages w…

Style TransferText Style Transfer

Multilingual pre-training with Language and Task Adaptation for Multilingual Text Style Transfer

2021-11-16 · ACL ARR November 2021 11 · Anonymous

We exploit the pre-trained seq2seq model mBART for multilingual text style transfer. Using machine translated data as well as gold aligned English sentences yields state-of-the-art results in the three target languages w…

Style TransferText Style Transfer