paper-with-me

홈 › Papers

Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs

2025-02-13 · Itai Mondshine, Tzuf Paz-Argaman, Reut Tsarfaty

Despite advances in the multilingual capabilities of Large Language Models (LLMs) across diverse tasks, English remains the dominant language for LLM research and development. So, when working with a different language, this has led to the widespread practice of pre-translation, i.e., translating the task prompt into English before inference. Selective pre-translation, a more surgical approach, focuses on translating specific prompt components. However, its current use is sporagic and lacks a systematic research foundation. Consequently, the optimal pre-translation strategy for various multilingual settings and tasks remains unclear. In this work, we aim to uncover the optimal setup for pre-translation by systematically assessing its use. Specifically, we view the prompt as a modular entity, composed of four functional parts: instruction, context, examples, and output, either of which could be translated or not. We evaluate pre-translation strategies across 35 languages covering both low and high-resource languages, on various tasks including Question Answering (QA), Natural Language Inference (NLI), Named Entity Recognition (NER), and Abstractive Summarization. Our experiments show the impact of factors as similarity to English, translation quality and the size of pre-trained data, on the model performance with pre-translation. We suggest practical guidelines for choosing optimal strategies in various multilingual settings.

📄 PDF Abstract BibTeX arXiv:2502.09331

Code (0)

등록된 구현이 없습니다.

Tasks

Abstractive Text Summarizationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)Natural Language InferenceNERQuestion AnsweringTranslation

Similar Papers 제목 키워드 기반

How and Where to Translate? The Impact of Translation Strategies in Cross-lingual LLM Prompting

2025-07-21 · Aman Gupta, Yingying Zhuang, Zhou Yu, Ziji Zhang 외 arxiv

Despite advances in the multilingual capabilities of Large Language Models (LLMs), their performance varies substantially across different languages and tasks. In multilingual retrieval-augmented generation (RAG)-based s…

Multilingual LLM Prompting Strategies for Medical English-Vietnamese Machine Translation

2025-09-19 · Nhu Vo, Nu-Uyen-Phuong Le, Dung D. Le, Massimo Piccardi 외 arxiv

Medical English-Vietnamese machine translation (En-Vi MT) is essential for healthcare access and communication in Vietnam, yet Vietnamese remains a low-resource and under-studied language. We systematically evaluate prom…

Machine Translation

Text2Cypher Across Languages: Evaluating Foundational Models Beyond English

2025-06-26 · Makbule Gulcin Ozsoy, William Tai

Recent advances in large language models have enabled natural language interfaces that translate user questions into database queries, such as Text2SQL, Text2SPARQL, and Text2Cypher. While these interfaces enhance databa…

AttributeText2Sparql

A Comparative Study of LLMs, NMT Models, and Their Combination in Persian-English Idiom Translation

2024-12-13 · Sara Rezaeimanesh, Faezeh Hosseini, Yadollah Yaghoobzadeh

Large language models (LLMs) have shown superior capabilities in translating figurative language compared to neural machine translation (NMT) systems. However, the impact of different prompting methods and LLM-NMT combin…

Machine TranslationNMTTranslation

The ADAPT System Description for the STAPLE 2020 English-to-Portuguese Translation Task

2020-07-01 · WS 2020 7 · Rejwanul Haque, Yasmin Moslem, Andy Way

This paper describes the ADAPT Centre{'}s submission to STAPLE (Simultaneous Translation and Paraphrase for Language Education) 2020, a shared task of the 4th Workshop on Neural Generation and Translation (WNGT), for the…

Machine TranslationSentenceTranslation