paper-with-me

Papers

Adversarial reconstruction for Multi-modal Machine Translation

2019-10-07 · Jean-Benoit Delbrouck, Stéphane Dupont

Even with the growing interest in problems at the intersection of Computer Vision and Natural Language, grounding (i.e. identifying) the components of a structured description in an image still remains a challenging task. This contribution aims to propose a model which learns grounding by reconstructing the visual features for the Multi-modal translation task. Previous works have partially investigated standard approaches such as regression methods to approximate the reconstruction of a visual input. In this paper, we propose a different and novel approach which learns grounding by adversarial feedback. To do so, we modulate our network following the recent promising adversarial architectures and evaluate how the adversarial response from a visual reconstruction as an auxiliary task helps the model in its learning. We report the highest scores in term of BLEU and METEOR metrics on the different datasets.

📄 PDF Abstract BibTeX arXiv:1910.02766

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Adversarial Evaluation of Multimodal Machine Translation

2018-10-01 · EMNLP 2018 10 · Desmond Elliott

The promise of combining language and vision in multimodal machine translation is that systems will produce better translations by leveraging the image data. However, the evidence surrounding whether the images are usefu…

Machine TranslationMultimodal Machine Translationtext similarityTranslation

Entity-level Cross-modal Learning Improves Multi-modal Machine Translation

2021-11-01 · Findings (EMNLP) 2021 11 · Xin Huang, Jiajun Zhang, Chengqing Zong

Multi-modal machine translation (MMT) aims at improving translation performance by incorporating visual information. Most of the studies leverage the visual information through integrating the global image features as au…

Machine TranslationRepresentation LearningTranslation

Deterministic Medical Image Translation via High-fidelity Brownian Bridges

2025-03-28 · Qisheng He, Nicholas Summerfield, Peiyong Wang, Carri Glide-Hurst 외

Recent studies have shown that diffusion models produce superior synthetic images when compared to Generative Adversarial Networks (GANs). However, their outputs are often non-deterministic and lack high fidelity to the …

Image Super-ResolutionSuper-ResolutionTranslation

Towards Multimodal Simultaneous Neural Machine Translation

2020-04-07 · WMT (EMNLP) 2020 11 · Aizhan Imankulova, Masahiro Kaneko, Tosho Hirasawa, Mamoru Komachi

Simultaneous translation involves translating a sentence before the speaker's utterance is completed in order to realize real-time understanding in multiple languages. This task is significantly more challenging than the…

Machine TranslationSentenceTranslation

Semi-Supervised Image-to-Image Translation

2019-01-24 · Manan Oza, Himanshu Vaghela, Sudhir Bagul

Image-to-image translation is a long-established and a difficult problem in computer vision. In this paper we propose an adversarial based model for image-to-image translation. The regular deep neural-network based metho…

Generative Adversarial NetworkImage SegmentationImage-to-Image TranslationMultimodal Unsupervised Image-To-Image Translation+4