paper-with-me

홈 › Papers

Do Multilingual Language Models Think Better in English?

2023-08-02 · Julen Etxaniz, Gorka Azkune, Aitor Soroa, Oier Lopez de Lacalle, Mikel Artetxe

Translate-test is a popular technique to improve the performance of multilingual language models. This approach works by translating the input into English using an external machine translation system, and running inference over the translated input. However, these improvements can be attributed to the use of a separate translation system, which is typically trained on large amounts of parallel data not seen by the language model. In this work, we introduce a new approach called self-translate, which overcomes the need of an external translation system by leveraging the few-shot translation capabilities of multilingual language models. Experiments over 5 tasks show that self-translate consistently outperforms direct inference, demonstrating that language models are unable to leverage their full multilingual potential when prompted in non-English languages. Our code is available at https://github.com/juletx/self-translate.

📄 PDF Abstract BibTeX arXiv:2308.01223

Code (1)

juletx/self-translate 공식 구현 pytorch

Tasks

Common Sense ReasoningCross-Lingual Natural Language InferenceCross-Lingual Paraphrase IdentificationLanguage ModelingLanguage ModellingMachine TranslationMath Word Problem SolvingNatural Language InferenceParaphrase IdentificationTranslation

Methods 이 논문이 사용한 방법론

BLOOM BLOOM is a decoder-only Transformer language model that was trained on the ROOTS corpus, a dataset comprising hundreds of sources in 46 natural and 13 programming languages…
LLaMA LLaMA is a collection of foundation language models ranging from 7B to 65B parameters. It is based on the transformer architecture with various improvements that were…

Similar Papers 제목 키워드 기반

Could Thinking Multilingually Empower LLM Reasoning?

2025-04-16 · Changjiang Gao, Xu Huang, Wenhao Zhu, ShuJian Huang 외

Previous work indicates that large language models exhibit a significant "English bias", i.e. they often perform better when tasks are presented in English. Interestingly, we have observed that using certain other langua…

Answer Selection

Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models

2024-11-02 · Hongyuan Lu, Zixuan Li, Wai Lam

As current training data for Large Language Models (LLMs) are dominated by English corpus, they are English-centric and they present impressive performance on English reasoning tasks.\footnote{This paper primarily studie…

GSM8KMath

Considerations for Multilingual Wikipedia Research

2022-04-05 · Isaac Johnson, Emily Lescak

English Wikipedia has long been an important data source for much research and natural language machine learning modeling. The growth of non-English language editions of Wikipedia, greater computational resources, and ca…

Language of Thought Shapes Output Diversity in Large Language Models

2026-01-16 · Shaoyang Xu, Wenxuan Zhang arxiv

Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the language used during model thinking-the language of thought-provides a novel an…

Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning

2025-10-08 · Xue Zhang, Yunlong Liang, Fandong Meng, Songming Zhang 외 arxiv

Large Reasoning Models (LRMs) have achieved remarkable performance on complex reasoning tasks by adopting the ``think-then-answer'' paradigm, which enhances both accuracy and interpretability. However, current LRMs exhib…

Reinforcement Learning