paper-with-me

홈 › Papers

Stronger Baselines for Trustable Results in Neural Machine Translation

2017-06-29 · WS 2017 8 · Michael Denkowski, Graham Neubig

Interest in neural machine translation has grown rapidly as its effectiveness has been demonstrated across language and data scenarios. New research regularly introduces architectural and algorithmic improvements that lead to significant gains over "vanilla" NMT implementations. However, these new techniques are rarely evaluated in the context of previously published techniques, specifically those that are widely used in state-of-theart production and shared-task systems. As a result, it is often difficult to determine whether improvements from research will carry over to systems deployed for real-world use. In this work, we recommend three specific methods that are relatively easy to implement and result in much stronger experimental systems. Beyond reporting significantly higher BLEU scores, we conduct an in-depth analysis of where improvements originate and what inherent weaknesses of basic NMT models are being addressed. We then compare the relative gains afforded by several other techniques proposed in the literature when starting with vanilla systems versus our stronger baselines, showing that experimental conclusions may change depending on the baseline chosen. This indicates that choosing a strong baseline is crucial for reporting reliable experimental results.

📄 PDF Abstract BibTeX arXiv:1706.09733

Code (1)

ijauregiCMCRC/ReWE_NMT pytorch

Tasks

Machine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

Approaching Neural Grammatical Error Correction as a Low-Resource Machine Translation Task

2018-04-16 · NAACL 2018 6 · Marcin Junczys-Dowmunt, Roman Grundkiewicz, Shubha Guha, Kenneth Heafield

Previously, neural methods in grammatical error correction (GEC) did not reach state-of-the-art results compared to phrase-based statistical machine translation (SMT) baselines. We demonstrate parallels between neural GE…

Domain AdaptationGrammatical Error CorrectionMachine TranslationTransfer Learning+1

A Unified Analytical Framework for Trustable Machine Learning and Automation Running with Blockchain

2019-03-21 · Tao Wang

Traditional machine learning algorithms use data from databases that are mutable, and therefore the data cannot be fully trusted. Also, the machine learning process is difficult to automate. This paper proposes building …

BIG-bench Machine Learning

Rule Learning as Machine Translation using the Atomic Knowledge Bank

2023-11-05 · Kristoffer Æsøy, Ana Ozaki

Machine learning models, and in particular language models, are being applied to various tasks that require reasoning. While such models are good at capturing patterns their ability to reason in a trustable and controlle…

Logical ReasoningMachine TranslationTranslation

From LLM to NMT: Advancing Low-Resource Machine Translation with Claude

2024-04-22 · Maxim Enis, Mark Hopkins

We show that Claude 3 Opus, a large language model (LLM) released by Anthropic in March 2024, exhibits stronger machine translation competence than other LLMs. Though we find evidence of data contamination with Claude on…

Knowledge DistillationLanguage ModelingLanguage ModellingLarge Language Model+3

Considerations for meaningful sign language machine translation based on glosses

2022-11-28 · Mathias Müller, Zifan Jiang, Amit Moryossef, Annette Rios 외

Automatic sign language processing is gaining popularity in Natural Language Processing (NLP) research (Yin et al., 2021). In machine translation (MT) in particular, sign language translation based on glosses is a promin…

Machine TranslationSign Language TranslationTranslation