paper-with-me

홈 › Papers

Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark

2025-08-28 · Chihiro Taguchi, Seng Mai, Keita Kurabe, Yusuke Sakai, Georgina Agyei, Soudabeh Eslami, David Chiang arxiv

Multilingual machine translation (MT) benchmarks play a central role in evaluating the capabilities of modern MT systems. Among them, the FLORES+ benchmark is widely used, offering English-to-many translation data for over 200 languages, curated with strict quality control protocols. However, we study data in four languages (Asante Twi, Japanese, Jinghpaw, and South Azerbaijani) and uncover critical shortcomings in the benchmark's suitability for truly multilingual evaluation. Human assessments reveal that many translations fall below the claimed 90% quality standard, and the annotators report that source sentences are often too domain-specific and culturally biased toward the English-speaking world. We further demonstrate that simple heuristics, such as copying named entities, can yield non-trivial BLEU scores, suggesting vulnerabilities in the evaluation protocol. Notably, we show that MT models trained on high-quality, naturalistic data perform poorly on FLORES+ while achieving significant gains on our domain-relevant evaluation set. Based on these findings, we advocate for multilingual MT benchmarks that use domain-general and culturally neutral source texts rely less on named entities, in order to better reflect real-world translation challenges.

📄 PDF Abstract BibTeX arXiv:2508.20511

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

The Less the Merrier? Investigating Language Representation in Multilingual Models

2023-10-20 · Hellina Hailu Nigatu, Atnafu Lambebo Tonja, Jugal Kalita

Multilingual Language Models offer a way to incorporate multiple languages in one model and utilize cross-language transfer learning to improve performance for different Natural Language Processing (NLP) tasks. Despite p…

named-entity-recognitionNamed Entity RecognitionText GenerationTransfer Learning

Letz Translate: Low-Resource Machine Translation for Luxembourgish

2023-03-02 · Yewei Song, Saad Ezzini, Jacques Klein, Tegawende Bissyande 외

Natural language processing of Low-Resource Languages (LRL) is often challenged by the lack of data. Therefore, achieving accurate machine translation (MT) in a low-resource environment is a real problem that requires pr…

Knowledge DistillationMachine TranslationTranslation

Extending Multilingual Machine Translation through Imitation Learning

2023-11-14 · Wen Lai, Viktor Hangya, Alexander Fraser

Despite the growing variety of languages supported by existing multilingual neural machine translation (MNMT) models, most of the world's languages are still being left behind. We aim to extend large-scale MNMT models to…

Imitation LearningMachine TranslationTranslation

Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages

2024-02-19 · Yuanchi Zhang, Yile Wang, Zijun Liu, Shuo Wang 외

While large language models (LLMs) have been pre-trained on multilingual corpora, their performance still lags behind in most languages compared to a few resource-rich languages. One common approach to mitigate this issu…

Transfer Learning

Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models

2026-08-15 · Varvara Arzt, Allan Hanbury, Terra Blevins arxiv

We systematically compare word order preferences in decoder-only language models across 192 artificial languages and typologically diverse natural languages. On artificial languages, models exhibit a left-branching prefe…