paper-with-me

Papers

Improving Multilingual Math Reasoning for African Languages

2025-05-26 · Odunayo Ogundepo, Akintunde Oladipo, Kelechi Ogueji, Esther Adenuga, David Ifeoluwa Adelani, Jimmy Lin

Researchers working on low-resource languages face persistent challenges due to limited data availability and restricted access to computational resources. Although most large language models (LLMs) are predominantly trained in high-resource languages, adapting them to low-resource contexts, particularly African languages, requires specialized techniques. Several strategies have emerged for adapting models to low-resource languages in todays LLM landscape, defined by multi-stage pre-training and post-training paradigms. However, the most effective approaches remain uncertain. This work systematically investigates which adaptation strategies yield the best performance when extending existing LLMs to African languages. We conduct extensive experiments and ablation studies to evaluate different combinations of data types (translated versus synthetically generated), training stages (pre-training versus post-training), and other model adaptation configurations. Our experiments focuses on mathematical reasoning tasks, using the Llama 3.1 model family as our base model.

📄 PDF Abstract BibTeX arXiv:2505.19848

Code (0)

등록된 구현이 없습니다.

Tasks

MathMathematical Reasoning

Methods 이 논문이 사용한 방법론

LLaMA LLaMA is a collection of foundation language models ranging from 7B to 65B parameters. It is based on the transformer architecture with various improvements that were…
BASE 설명 없음

Similar Papers 제목 키워드 기반

AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages

2026-01-10 · Hao Yu, Tianyi Xu, Michael A. Hedderich, Wassim Hamidouche 외 arxiv

Large language models (LLMs) are increasingly multilingual, yet open models continue to underperform relative to proprietary systems, with the gap most pronounced for African languages. Continued pre-training (CPT) offer…

Mathematical Reasoning

Tencent's Multilingual Machine Translation System for WMT22 Large-Scale African Languages

2022-10-18 · Wenxiang Jiao, Zhaopeng Tu, Jiarui Li, Wenxuan Wang 외

This paper describes Tencent's multilingual machine translation systems for the WMT22 shared task on Large-Scale Machine Translation Evaluation for African Languages. We participated in the $\mathbf{constrained}$ transla…

Data AugmentationMachine TranslationTranslation

Crosslingual On-Policy Self-Distillation for Multilingual Reasoning

2026-05-10 · Yihong Liu, Raoyuan Zhao, Michael A. Hedderich, Hinrich Schütze arxiv

Large language models (LLMs) have achieved remarkable progress in mathematical reasoning, but this ability is not equally accessible across languages. Especially low-resource languages exhibit much lower reasoning perfor…

Reinforcement LearningMathematical Reasoning

MMTAfrica: Multilingual Machine Translation for African Languages

2022-04-08 · WMT (EMNLP) 2021 11 · Chris C. Emezue, Bonaventure F. P. Dossou

In this paper, we focus on the task of multilingual machine translation for African languages and describe our contribution in the 2021 WMT Shared Task: Large-Scale Multilingual Machine Translation. We introduce MMTAfric…

Machine TranslationTranslation

Preparing the Vuk'uzenzele and ZA-gov-multilingual South African multilingual corpora

2023-03-07 · Richard Lastrucci, Isheanesu Dzingirai, Jenalea Rajab, Andani Madodonga 외

This paper introduces two multilingual government themed corpora in various South African languages. The corpora were collected by gathering the South African Government newspaper (Vuk'uzenzele), as well as South African…

Language ModelingLanguage ModellingMachine TranslationNMT+1