paper-with-me

홈 › Papers

Incorporating Textual Evidence in Visual Storytelling

2019-11-21 · WS 2019 11 · Tianyi Li, Sujian Li

Previous work on visual storytelling mainly focused on exploring image sequence as evidence for storytelling and neglected textual evidence for guiding story generation. Motivated by human storytelling process which recalls stories for familiar images, we exploit textual evidence from similar images to help generate coherent and meaningful stories. To pick the images which may provide textual experience, we propose a two-step ranking method based on image object recognition techniques. To utilize textual information, we design an extended Seq2Seq model with two-channel encoder and attention. Experiments on the VIST dataset show that our method outperforms state-of-the-art baseline models without heavy engineering.

📄 PDF Abstract BibTeX arXiv:1911.09334

Code (0)

등록된 구현이 없습니다.

Tasks

Object RecognitionStory GenerationVisual Storytelling

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…
Seq2Seq Seq2Seq, or Sequence To Sequence, is a model used in sequence prediction tasks, such as language modelling and machine translation. The idea is to use one…

Similar Papers 제목 키워드 기반

Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning

2024-08-12 · Yingjin Song, Denis Paperno, Albert Gatt

Visual storytelling systems generate multi-sentence stories from image sequences. In this task, capturing contextual information and bridging visual variation bring additional challenges. We propose a simple yet effectiv…

Contrastive LearningInformativenessSentenceVisual Storytelling

ContextualStory: Consistent Visual Storytelling with Spatially-Enhanced and Storyline Context

2024-07-13 · Sixiao Zheng, Yanwei Fu

Visual storytelling involves generating a sequence of coherent frames from a textual storyline while maintaining consistency in characters and scenes. Existing autoregressive methods, which rely on previous frame-sentenc…

Image GenerationStory ContinuationStory VisualizationText-to-Image Generation+1

AESOP: Abstract Encoding of Stories, Objects, and Pictures

2021-01-01 · ICCV 2021 10 · Hareesh Ravi, Kushal Kafle, Scott Cohen, Jonathan Brandt 외

Visual storytelling and story comprehension are uniquely human skills that play a central role in how we learn about and experience the world. Despite remarkable progress in recent years in synthesis of visual and te…

Story CompletionVisual Storytelling

VIST-GPT: Ushering in the Era of Visual Storytelling with LLMs?

2025-04-27 · Mohamed Gado, Towhid Taliee, Muhammad Memon, Dmitry Ignatov 외

Visual storytelling is an interdisciplinary field combining computer vision and natural language processing to generate cohesive narratives from sequences of images. This paper presents a novel approach that leverages re…

Visual GroundingVisual Storytelling

A Pipeline for Creative Visual Storytelling

2018-07-21 · WS 2018 6 · Stephanie M. Lukin, Reginald Hobbs, Clare R. Voss

Computational visual storytelling produces a textual description of events and interpretations depicted in a sequence of images. These texts are made possible by advances and cross-disciplinary approaches in natural lang…

Visual Storytelling