paper-with-me

홈 › Papers

Movies2Scenes: Using Movie Metadata to Learn Scene Representation

2022-02-22 · CVPR 2023 1 · Shixing Chen, Chun-Hao Liu, Xiang Hao, Xiaohan Nie, Maxim Arap, Raffay Hamid

Understanding scenes in movies is crucial for a variety of applications such as video moderation, search, and recommendation. However, labeling individual scenes is a time-consuming process. In contrast, movie level metadata (e.g., genre, synopsis, etc.) regularly gets produced as part of the film production process, and is therefore significantly more commonly available. In this work, we propose a novel contrastive learning approach that uses movie metadata to learn a general-purpose scene representation. Specifically, we use movie metadata to define a measure of movie similarity, and use it during contrastive learning to limit our search for positive scene-pairs to only the movies that are considered similar to each other. Our learned scene representation consistently outperforms existing state-of-the-art methods on a diverse set of tasks evaluated using multiple benchmark datasets. Notably, our learned representation offers an average improvement of 7.9% on the seven classification tasks and 9.7% improvement on the two regression tasks in LVU dataset. Furthermore, using a newly collected movie dataset, we present comparative results of our scene representation on a set of video moderation tasks to demonstrate its generalizability on previously less explored tasks.

📄 PDF Abstract BibTeX arXiv:2202.10650

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningScene Understanding

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Condensed Movies: Story Based Retrieval with Contextual Embeddings

2020-05-08 · Max Bain, Arsha Nagrani, Andrew Brown, Andrew Zisserman

Our objective in this work is long range understanding of the narrative structure of movies. Instead of considering the entire movie, we propose to learn from the `key scenes' of the movie, providing a condensed look at …

RetrievalText to Video RetrievalVideo Retrieval

A Local-to-Global Approach to Multi-modal Movie Scene Segmentation

2020-04-06 · CVPR 2020 6 · Anyi Rao, Linning Xu, Yu Xiong, Guodong Xu 외

Scene, as the crucial unit of storytelling in movies, contains complex activities of actors and their interactions in a physical environment. Identifying the composition of scenes serves as a critical step towards semant…

Action RecognitionScene SegmentationSegmentation

MovieCLIP: Visual Scene Recognition in Movies

2022-10-20 · Digbalay Bose, Rajat Hebbar, Krishna Somandepalli, Haoyang Zhang 외

Longform media such as movies have complex narrative structures, with events spanning a rich variety of ambient visual scenes. Domain specific challenges associated with visual scenes in movies include transitions, perso…

Genre classificationScene Recognition

Moviescope: Large-scale Analysis of Movies using Multiple Modalities

2019-08-08 · Paola Cascante-Bonilla, Kalpathy Sitaraman, Mengjia Luo, Vicente Ordonez

Film media is a rich form of artistic expression. Unlike photography, and short videos, movies contain a storyline that is deliberately complex and intricate in order to engage its audience. In this paper we present a la…

Indian Regional Movie Dataset for Recommender Systems

2018-01-07 · Prerna Agarwal, Richa Verma, Angshul Majumdar

Indian regional movie dataset is the first database of regional Indian movies, users and their ratings. It consists of movies belonging to 18 different Indian regional languages and metadata of users with varying demogra…

Collaborative Filteringcompressed sensingDiversityMatrix Completion+1