paper-with-me

Papers

Keep it Consistent: Topic-Aware Storytelling from an Image Stream via Iterative Multi-agent Communication

2019-11-11 · COLING 2020 8 · Ruize Wang, Zhongyu Wei, Ying Cheng, Piji Li, Haijun Shan, Ji Zhang, Qi Zhang, Xuanjing Huang

Visual storytelling aims to generate a narrative paragraph from a sequence of images automatically. Existing approaches construct text description independently for each image and roughly concatenate them as a story, which leads to the problem of generating semantically incoherent content. In this paper, we propose a new way for visual storytelling by introducing a topic description task to detect the global semantic context of an image stream. A story is then constructed with the guidance of the topic description. In order to combine the two generation tasks, we propose a multi-agent communication framework that regards the topic description generator and the story generator as two agents and learn them simultaneously via iterative updating mechanism. We validate our approach on VIST dataset, where quantitative results, ablations, and human evaluation demonstrate our method's good ability in generating stories with higher quality compared to state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:1911.04192

Code (0)

등록된 구현이 없습니다.

Tasks

Image CaptioningQuestion GenerationVisual Storytelling

Similar Papers 제목 키워드 기반

TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling

2024-03-18 · Weiran Chen, Xin Li, Jiaqi Su, Guiqian Zhu 외

As a cross-modal task, visual storytelling aims to generate a story for an ordered image sequence automatically. Different from the image captioning task, visual storytelling requires not only modeling the relationships …

Image CaptioningVisual Storytelling

MementoEmbed and Raintale for Web Archive Storytelling

2020-08-01 · Shawn M. Jones, Martin Klein, Michele C. Weigle, Michael L. Nelson

For traditional library collections, archivists can select a representative sample from a collection and display it in a featured physical or digital library space. Web archive collections may consist of thousands of arc…

Coherent Visual Storytelling via Parallel Top-Down Visual and Topic Attention

2022-08-17 · IEEE Transactions on Circuits and Systems for Video Technology 2022 8 · Jinjing Gu, Hanli Wang

Visual storytelling aims at producing a narrative paragraph for a given photo album automatically. It introduces more new challenges than individual image paragraph descriptions, mainly due to the difficulty in preservin…

DiversitySentenceText GenerationVisual Storytelling

Visual Storytelling with Hierarchical BERT Semantic Guidance

2022-01-10 · ACM Multimedia Asia 2022 1 · Ruichao Fan, Hanli Wang, Jinjing Gu, and Xianhui Liu

Visual storytelling, which aims at automatically producing a narrative paragraph for photo album, remains quite challenging due to the complexity and diversity of photo album content. In addition, open-domain photo album…

SentenceText GenerationVisual Storytelling

Knowledge-enriched Attention Network with Group-wise Semantic for Visual Storytelling

2022-03-10 · Tengpeng Li, Hanli Wang, Bin He, Chang Wen Chen

As a technically challenging topic, visual storytelling aims at generating an imaginary and coherent story with narrative multi-sentences from a group of relevant images. Existing methods often generate direct and rigid …

DecoderStory GenerationVisual Storytelling