paper-with-me

홈 › Papers

Scene-driven Retrieval in Edited Videos using Aesthetic and Semantic Deep Features

2016-04-09 · Lorenzo Baraldi, Costantino Grana, Rita Cucchiara

This paper presents a novel retrieval pipeline for video collections, which aims to retrieve the most significant parts of an edited video for a given query, and represent them with thumbnails which are at the same time semantically meaningful and aesthetically remarkable. Videos are first segmented into coherent and story-telling scenes, then a retrieval algorithm based on deep learning is proposed to retrieve the most significant scenes for a textual query. A ranking strategy based on deep features is finally used to tackle the problem of visualizing the best thumbnail. Qualitative and quantitative experiments are conducted on a collection of edited videos to demonstrate the effectiveness of our approach.

📄 PDF Abstract BibTeX arXiv:1604.02546

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval

Similar Papers 제목 키워드 기반

GAZED- Gaze-guided Cinematic Editing of Wide-Angle Monocular Video Recordings

2020-10-22 · K L Bhanu Moorthy, Moneish Kumar, Ramanathan Subramaniam, Vineet Gandhi

We present GAZED- eye GAZe-guided EDiting for videos captured by a solitary, static, wide-angle and high-resolution camera. Eye-gaze has been effectively employed in computational applications as a cue to capture interes…

validVideo Editing

VE-Bench: Subjective-Aligned Benchmark Suite for Text-Driven Video Editing Quality Assessment

2024-08-21 · Shangkun Sun, Xiaoyu Liang, Songlin Fan, Wenxu Gao 외

Text-driven video editing has recently experienced rapid development. Despite this, evaluating edited videos remains a considerable challenge. Current metrics tend to fail to align with human perceptions, and effective q…

Video AlignmentVideo EditingVideo Quality AssessmentVisual Question Answering (VQA)

A gaze driven fast-forward method for first-person videos

2020-06-10 · Alan Carvalho Neves, Michel Melo Silva, Mario Fernando Montenegro Campos, Erickson Rangel Nascimento

The growing data sharing and life-logging cultures are driving an unprecedented increase in the amount of unedited First-Person Videos. In this paper, we address the problem of accessing relevant information in First-Per…

Aesthetics Driven Autonomous Time-Lapse Photography Generation by Virtual and Real Robots

2022-08-22 · Xiaobo Gao, Qi Kuang, Xin Jin, Bin Zhou 외

Time-lapse photography is employed in movies and promotional films because it can reflect the passage of time in a short time and strengthen the visual attraction. However, since it takes a long time and requires the sta…

Recognizing and Presenting the Storytelling Video Structure with Deep Multimodal Networks

2016-10-05 · Lorenzo Baraldi, Costantino Grana, Rita Cucchiara

This paper presents a novel approach for temporal and semantic segmentation of edited videos into meaningful segments, from the point of view of the storytelling structure. The objective is to decompose a long video into…

Change DetectionRetrievalSemantic Segmentation