paper-with-me

홈 › Papers

Improving Image Captioning with Conditional Generative Adversarial Nets

2018-05-18 · Chen Chen, Shuai Mu, Wanpeng Xiao, Zexiong Ye, Liesi Wu, Qi Ju

In this paper, we propose a novel conditional-generative-adversarial-nets-based image captioning framework as an extension of traditional reinforcement-learning (RL)-based encoder-decoder architecture. To deal with the inconsistent evaluation problem among different objective language metrics, we are motivated to design some "discriminator" networks to automatically and progressively determine whether generated caption is human described or machine generated. Two kinds of discriminator architectures (CNN and RNN-based structures) are introduced since each has its own advantages. The proposed algorithm is generic so that it can enhance any existing RL-based image captioning framework and we show that the conventional RL training method is just a special case of our approach. Empirically, we show consistent improvements over all language evaluation metrics for different state-of-the-art image captioning models. In addition, the well-trained discriminators can also be viewed as objective image captioning evaluators

📄 PDF Abstract BibTeX arXiv:1805.07112

Code (1)

Anjaney1999/image-captioning-seqgan pytorch

Tasks

DecoderImage CaptioningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

A Thorough Review on Recent Deep Learning Methodologies for Image Captioning

2021-07-28 · Ahmed Elhagry, Karima Kadaoui

Image Captioning is a task that combines computer vision and natural language processing, where it aims to generate descriptive legends for images. It is a two-fold process relying on accurate image understanding and cor…

Caption GenerationDescriptiveImage CaptioningMeta-Learning

Conditional Generative Adversarial Nets

2014-11-06 · Mehdi Mirza, Simon Osindero

Generative Adversarial Nets [8] were recently introduced as a novel way to train generative models. In this work we introduce the conditional version of generative adversarial nets, which can be constructed by simply fee…

DescriptiveHuman action generation

Face Aging with Contextual Generative Adversarial Nets

2018-02-01 · Si Liu, Yao Sun, Defa Zhu, Renda Bao 외

Face aging, which renders aging faces for an input face, has attracted extensive attention in the multimedia research. Recently, several conditional Generative Adversarial Nets (GANs) based methods have achieved great su…

Face Verification

Can adversarial training learn image captioning ?

2019-10-31 · Jean-Benoit Delbrouck, Bastien Vanderplaetse, Stéphane Dupont

Recently, generative adversarial networks (GAN) have gathered a lot of interest. Their efficiency in generating unseen samples of high quality, especially images, has improved over the years. In the field of Natural Lang…

Image CaptioningText Generation

Triple Generative Adversarial Nets

2017-03-07 · NeurIPS 2017 12 · Chongxuan Li, Kun Xu, Jun Zhu, Bo Zhang

Generative Adversarial Nets (GANs) have shown promise in image generation and semi-supervised learning (SSL). However, existing GANs in SSL have two problems: (1) the generator and the discriminator (i.e. the classifier)…

Image Generation