paper-with-me

홈 › Papers

BackTranslation2.0 -- A Linguistically Motivated Metric to Assess Sign Language Production

2026-06-27 · Oliver Cory, Maksym Ivashechkin, Karahan Sahin, Oline Ranum, Jianhe Low, Edward Fish, Anton Pelykh, Ozge Mercanoglu Sincan, Richard Bowden arxiv

Sign Languages (SLs) are the primary means of communication for millions of deaf individuals, yet existing evaluation metrics for generated SL remain simplistic and poorly aligned with human judgements. We introduce BackTranslation2.0, a linguistically grounded evaluation metric for text-to-sign translation that moves beyond naïve backtranslation. Our approach adopts an agentic framework in which a deterministic pipeline orchestrates a suite of specialised tools to assess four scoring dimensions - grammatical correctness, phonological accuracy, motion fluency, and generation fidelity - aligned with human rater assessments. Tool outputs are not treated independently: a set of large language model (LLM)-based cross-referential comparison modules evaluates consistency across tools and checks outputs against linguistic expectations, enabling structured reasoning over grammatical, phonological, and motion-level evidence. Final dimension scores are computed through deterministic weighted formulas over validated tool outputs. To validate BackTranslation2.0, we introduce and evaluate on a British Sign Language (BSL) dataset rated in a human rater study across the same quality dimensions, following a protocol developed in collaboration between linguists and deaf experts, benchmarking against six baseline metrics. Our method demonstrates strong correlation with human judgements across all dimensions, providing a more comprehensive, interpretable, and linguistically principled evaluation framework for sign language production systems.

📄 PDF Abstract BibTeX arXiv:2606.28673

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Data Augmentation for Low-Resource Named Entity Recognition Using Backtranslation

2021-08-26 · ICON 2021 12 · Usama Yaseen, Stefan Langer

The state of art natural language processing systems relies on sizable training datasets to achieve high performance. Lack of such datasets in the specialized low resource domains lead to suboptimal performance. In this …

Data AugmentationLow Resource Named Entity Recognitionnamed-entity-recognitionNamed Entity Recognition+1

Linguistic Features for Readability Assessment

2020-05-30 · WS 2020 7 · Tovly Deutsch, Masoud Jasbi, Stuart Shieber

Readability assessment aims to automatically classify text by the level appropriate for learning readers. Traditional approaches to this task utilize a variety of linguistically motivated features paired with simple mach…

Deep LearningText Classification

VERTa: a Linguistically-motivated Metric at the WMT15 Metrics Task

2015-09-01 · WS 2015 9 · Elisabet Comelles, Jordi Atserias
Language ModellingMachine Translation

How does the pre-training objective affect what large language models learn about linguistic properties?

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Several pre-training objectives, such as masked language modeling (MLM), have been proposed to pre-train language models (e.g. BERT) with the aim of learning better language representations. However, to the best of our k…

Language ModelingLanguage ModellingMasked Language Modeling

LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation

2026-04-20 · Samar M. Magdy, Fakhraddin Alwajih, Abdellah El Mekki, Wesam El-Sayed 외 arxiv

Existing MT evaluation frameworks, including automatic metrics and human evaluation schemes such as Multidimensional Quality Metrics (MQM), are largely language-agnostic. However, they often fail to capture dialect- and …

Machine Translation