paper-with-me

홈 › Papers

ENGINE: Energy-Based Inference Networks for Non-Autoregressive Machine Translation

2020-05-02 · ACL 2020 6 · Lifu Tu, Richard Yuanzhe Pang, Sam Wiseman, Kevin Gimpel

We propose to train a non-autoregressive machine translation model to minimize the energy defined by a pretrained autoregressive model. In particular, we view our non-autoregressive translation system as an inference network (Tu and Gimpel, 2018) trained to minimize the autoregressive teacher energy. This contrasts with the popular approach of training a non-autoregressive model on a distilled corpus consisting of the beam-searched outputs of such a teacher model. Our approach, which we call ENGINE (ENerGy-based Inference NEtworks), achieves state-of-the-art non-autoregressive results on the IWSLT 2014 DE-EN and WMT 2016 RO-EN datasets, approaching the performance of autoregressive models.

📄 PDF Abstract BibTeX arXiv:2005.00850

Code (1)

lifu-tu/ENGINE 공식 구현 pytorch

Tasks

de-enMachine TranslationTranslation

Similar Papers 제목 키워드 기반

The RoyalFlush System for the WMT 2022 Efficiency Task

2022-12-03 · Bo Qin, Aixin Jia, Qiang Wang, Jianning Lu 외

This paper describes the submission of the RoyalFlush neural machine translation system for the WMT 2022 translation efficiency task. Unlike the commonly used autoregressive translation system, we adopted a two-stage tra…

DecoderGPUKnowledge DistillationMachine Translation+1

Latent-Variable Non-Autoregressive Neural Machine Translation with Deterministic Inference Using a Delta Posterior

2019-08-20 · Raphael Shu, Jason Lee, Hideki Nakayama, Kyunghyun Cho

Although neural machine translation models reached high translation quality, the autoregressive nature makes inference difficult to parallelize and leads to high translation latency. Inspired by recent refinement-based a…

Machine TranslationTranslation

Non-Autoregressive Translation with Layer-Wise Prediction and Deep Supervision

2021-10-14 · Chenyang Huang, Hao Zhou, Osmar R. Zaïane, Lili Mou 외

How do we perform efficient inference while retaining high translation quality? Existing neural machine translation models, such as Transformer, achieve high performance, but they decode words one by one, which is ineffi…

Machine TranslationTranslation

Imitation Learning for Non-Autoregressive Neural Machine Translation

2019-06-05 · ACL 2019 7 · Bingzhen Wei, Mingxuan Wang, Hao Zhou, Junyang Lin 외

Non-autoregressive translation models (NAT) have achieved impressive inference speedup. A potential issue of the existing NAT algorithms, however, is that the decoding is conducted in parallel, without directly consideri…

Imitation LearningMachine TranslationSentenceTranslation

Hint-Based Training for Non-Autoregressive Machine Translation

2019-09-15 · IJCNLP 2019 11 · Zhuohan Li, Zi Lin, Di He, Fei Tian 외

Due to the unparallelizable nature of the autoregressive factorization, AutoRegressive Translation (ART) models have to generate tokens sequentially during decoding and thus suffer from high inference latency. Non-AutoRe…

de-enMachine TranslationTranslation