paper-with-me

Papers

Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling

2021-02-05 · Hong Chen, Yifei HUANG, Hiroya Takamura, Hideki Nakayama

Visual storytelling is a task of generating relevant and interesting stories for given image sequences. In this work we aim at increasing the diversity of the generated stories while preserving the informative content from the images. We propose to foster the diversity and informativeness of a generated story by using a concept selection module that suggests a set of concept candidates. Then, we utilize a large scale pre-trained model to convert concepts and images into full stories. To enrich the candidate concepts, a commonsense knowledge graph is created for each image sequence from which the concept candidates are proposed. To obtain appropriate concepts from the graph, we propose two novel modules that consider the correlation among candidate concepts and the image-concept correlation. Extensive automatic and human evaluation results demonstrate that our model can produce reasonable concepts. This enables our model to outperform the previous models by a large margin on the diversity and informativeness of the story, while retaining the relevance of the story to the image sequence.

📄 PDF Abstract BibTeX arXiv:2102.02963

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityInformativenessVisual Storytelling

Similar Papers 제목 키워드 기반

ASER: Towards Large-scale Commonsense Knowledge Acquisition via Higher-order Selectional Preference over Eventualities

2021-04-05 · Hongming Zhang, Xin Liu, Haojie Pan, Haowen Ke 외

Commonsense knowledge acquisition and reasoning have long been a core artificial intelligence problem. However, in the past, there has been a lack of scalable methods to collect commonsense knowledge. In this paper, we p…

Discourse Parsing

Knowledgeable Storyteller: A Commonsense-Driven Generative Model for Visual Storytelling

2019-05-04 · IJCAI 2019 2019 5 · Pengcheng Yang, Fuli Luo, Peng Chen, Lei LI 외

The visual storytelling (VST) task aims at generating a reasonable and coherent paragraph-level story with the image stream as input. Different from caption that is a direct and literal description of image content, the …

AI AgentKnowledge GraphsSemantic SimilaritySemantic Textual Similarity+2

Visual Commonsense-aware Representation Network for Video Captioning

2022-11-17 · Pengpeng Zeng, Haonan Zhang, Lianli Gao, Xiangpeng Li 외

Generating consecutive descriptions for videos, i.e., Video Captioning, requires taking full advantage of visual representation along with the generation process. Existing video captioning methods focus on making an expl…

Caption GenerationQuestion AnsweringVideo CaptioningVideo Question Answering

Acquiring and Modelling Abstract Commonsense Knowledge via Conceptualization

2022-06-03 · Mutian He, Tianqing Fang, Weiqi Wang, Yangqiu Song

Conceptualization, or viewing entities and situations as instances of abstract concepts in mind and making inferences based on that, is a vital component in human intelligence for commonsense reasoning. Despite recent pr…

Knowledge Graphs

COMET: Commonsense Transformers for Automatic Knowledge Graph Construction

2019-06-12 · ACL 2019 7 · Antoine Bosselut, Hannah Rashkin, Maarten Sap, Chaitanya Malaviya 외

We present the first comprehensive study on automatic knowledge base construction for two prevalent commonsense knowledge graphs: ATOMIC (Sap et al., 2019) and ConceptNet (Speer et al., 2017). Contrary to many convention…

graph constructionKnowledge Base ConstructionKnowledge Graphs