paper-with-me

Papers

Dual-CNN: A Convolutional language decoder for paragraph image captioning

2020-02-14 · Neurocomputing 2020 2 · Ruifan Li, Haoyun Liang, Yihui Shi, Fangxiang Feng, Xiaojie Wang

Abstract The task of paragraph image captioning aims to generate a coherent paragraph describing a given image. However, due to their limited ability to capture long-term dependency, recurrent neural network or long-short term memory based decoders could hardly generate satisfactory textual descriptions with a long paragraph. In addition, the training inefficiency in the sequential decoders is significantly observed. Motivated by the advantage of convolutional neural network (i.e., CNN), in this paper, we propose a Dual-CNN decoder with long-term memory ability and parallel computation, which can produce a semantically coherent paragraph for an image. Our Dual-CNN model is evaluated on the Stanford image-paragraph dataset. Extensive experiments demonstrate that our Dual-CNN achieves comparable results compared with state-of-the-art models. Furthermore, the diversity and coherence of generated paragraphs are analyzed to show the superiority of our approach.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDiversityImage CaptioningImage Paragraph Captioning

Similar Papers 제목 키워드 기반

Improving Diversity and Reducing Redundancy in Paragraph Captions

2020-07-19 · International Joint Conference on Neural Networks (IJCNN) 2020 7 · Kanani, Chandresh S., Sriparna Saha, and Pushpak Bhattacharyya

The purpose of an image paragraph captioning model is to produce detailed descriptions of the source images. Generally, paragraph captioning models use encoder-decoder based architectures similar to the standard image…

DecoderDense CaptioningDiversityImage Captioning+1

Byte-Level Recursive Convolutional Auto-Encoder for Text

2018-02-06 · ICLR 2018 1 · Xiang Zhang, Yann Lecun

This article proposes to auto-encode text at byte-level using convolutional networks with a recursive architecture. The motivation is to explore whether it is possible to have scalable and homogeneous text generation at …

DecoderText Generation

Text Generation with Diffusion Language Models: A Pre-training Approach with Continuous Paragraph Denoise

2022-12-22 · Zhenghao Lin, Yeyun Gong, Yelong Shen, Tong Wu 외

In this paper, we introduce a novel dIffusion language modEl pre-training framework for text generation, which we call GENIE. GENIE is a large-scale pretrained diffusion language model that consists of an encoder and a d…

DecoderDenoisingLanguage ModelingLanguage Modelling+1

Hierarchical Scene Graph Encoder-Decoder for Image Paragraph Captioning

2020-10-12 · ACM International Conference on Multimedia 2020 10 · Yang, Xu, Chongyang Gao, Hanwang Zhang 외

When we humans tell a long paragraph about an image, we usually first implicitly compose a mental “script” and then comply with it to generate the paragraph. Inspired by this, we render the modern encoder-decoder base…

DecoderImage Paragraph CaptioningSentence

Convolutional Auto-encoding of Sentence Topics for Image Paragraph Generation

2019-08-01 · Jing Wang, Yingwei Pan, Ting Yao, Jinhui Tang 외

Image paragraph generation is the task of producing a coherent story (usually a paragraph) that describes the visual content of an image. The problem nevertheless is not trivial especially when there are multiple descrip…

DescriptiveImage Paragraph CaptioningSentencevalid