paper-with-me

Papers

MetricX-24: The Google Submission to the WMT 2024 Metrics Shared Task

2024-10-04 · Juraj Juraska, Daniel Deutsch, Mara Finkelstein, Markus Freitag

In this paper, we present the MetricX-24 submissions to the WMT24 Metrics Shared Task and provide details on the improvements we made over the previous version of MetricX. Our primary submission is a hybrid reference-based/-free metric, which can score a translation irrespective of whether it is given the source segment, the reference, or both. The metric is trained on previous WMT data in a two-stage fashion, first on the DA ratings only, then on a mixture of MQM and DA ratings. The training set in both stages is augmented with synthetic examples that we created to make the metric more robust to several common failure modes, such as fluent but unrelated translation, or undertranslation. We demonstrate the benefits of the individual modifications via an ablation study, and show a significant performance increase over MetricX-23 on the WMT23 MQM ratings, as well as our new synthetic challenge set.

📄 PDF Abstract BibTeX arXiv:2410.03983

Code (1)

google-research/metricx 공식 구현 pytorch

Tasks

Translation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

MetricX-25 and GemSpanEval: Google Translate Submissions to the WMT25 Evaluation Shared Task

2025-10-28 · Juraj Juraska, Tobias Domhan, Mara Finkelstein, Tetsuji Nakagawa 외 arxiv

In this paper, we present our submissions to the unified WMT25 Translation Evaluation Shared Task. For the Quality Score Prediction subtask, we create a new generation of MetricX with improvements in the input format and…

HydraQE: OSU's Submission for the IWSLT 2026 Speech Translation Metrics Shared Task

2026-06-07 · Kevin Krahn, Eric Fosler-Lussier arxiv

We present HydraQE, our contribution to the IWSLT 2026 Speech Translation Metrics shared task. HydraQE is an end-to-end, reference-free quality estimation (QE) system for speech translation built on a Qwen3-ASR backbone,…

Machine Translation

UHH Submission to the WMT17 Metrics Shared Task

2017-09-01 · WS 2017 9 · Melania Duma, Wolfgang Menzel
Machine Translation

MTEQA at WMT21 Metrics Shared Task

2021-11-01 · WMT (EMNLP) 2021 11 · Mateusz Krubiński, Erfan Ghadery, Marie-Francine Moens, Pavel Pecina

In this paper, we describe our submission to the WMT 2021 Metrics Shared Task. We use the automatically-generated questions and answers to evaluate the quality of Machine Translation (MT) systems. Our submission builds u…

Machine TranslationTranslation

Shared Task on Evaluating Accuracy

2020-12-01 · INLG (ACL) 2020 12 · Ehud Reiter, Craig Thomson

We propose a shared task on methodologies and algorithms for evaluating the accuracy of generated texts, specifically summaries of basketball games produced from basketball box score and other game data. We welcome submi…