paper-with-me

홈 › Papers

Movienet: A Movie Multilayer Network Model using Visual and Textual Semantic Cues

2019-10-18 · Youssef Mourchid, Benjamin Renoust, Olivier Roupin, Le Van, Hocine Cherifi, Mohammed El Hassouni

Discovering content and stories in movies is one of the most important concepts in multimedia content research studies. Network models have proven to be an efficient choice for this purpose. When an audience watches a movie, they usually compare the characters and the relationships between them. For this reason, most of the models developed so far are based on social networks analysis. They focus essentially on the characters at play. By analyzing characters' interactions, we can obtain a broad picture of the narration's content. Other works have proposed to exploit semantic elements such as scenes, dialogues, etc. However, they are always captured from a single facet. Motivated by these limitations, we introduce in this work a multilayer network model to capture the narration of a movie based on its script, its subtitles, and the movie content. After introducing the model and the extraction process from the raw data, we perform a comparative analysis of the whole 6-movie cycle of the Star Wars saga. Results demonstrate the effectiveness of the proposed framework for video content representation and analysis.

📄 PDF Abstract BibTeX arXiv:1910.09368

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MovieNet: A Holistic Dataset for Movie Understanding

2020-07-21 · ECCV 2020 8 · Qingqiu Huang, Yu Xiong, Anyi Rao, Jiaze Wang 외

Recent years have seen remarkable advances in visual understanding. However, how to understand a story-based long video with artistic styles, e.g. movie, remains challenging. In this paper, we introduce MovieNet -- a hol…

Video Understanding

MoviePuzzle: Visual Narrative Reasoning through Multimodal Order Learning

2023-06-04 · Jianghui Wang, Yuxuan Wang, Dongyan Zhao, Zilong Zheng

We introduce MoviePuzzle, a novel challenge that targets visual narrative reasoning and holistic movie understanding. Despite the notable progress that has been witnessed in the realm of video understanding, most prior w…

BenchmarkingContrastive LearningVideo Understanding

TeViS:Translating Text Synopses to Video Storyboards

2022-12-31 · Xu Gu, Yuchong Sun, Feiyue Ni, ShiZhe Chen 외

A video storyboard is a roadmap for video creation which consists of shot-by-shot images to visualize key plots in a text synopsis. Creating video storyboards, however, remains challenging which not only requires cross-m…

Language ModellingQuantization

Scene-VLM: Multimodal Video Scene Segmentation via Vision-Language Models

2025-12-25 · Nimrod Berman, Adam Botach, Emanuel Ben-Baruch, Shunit Haviv Hakimi 외 arxiv

Segmenting long-form videos into semantically coherent scenes is a fundamental task in large-scale video understanding. Existing encoder-based methods are limited by visual-centric biases, classify each shot in isolation…

Multimodal ReasoningScene Segmentation

Movie Trailer Genre Classification Using Multimodal Pretrained Features

2024-10-11 · Serkan Sulun, Paula Viana, Matthew E. P. Davies

We introduce a novel method for movie genre classification, capitalizing on a diverse set of readily accessible pretrained models. These models extract high-level features related to visual scenery, objects, characters, …

ClassificationGenre classification