Video Summarization Using Fully Convolutional Sequence Networks
This paper addresses the problem of video summarization. Given an input video, the goal is to select a subset of the frames to create a summary video that optimally captures the important information of the input video. With the large amount of videos available online, video summarization provides a useful tool that assists video search, retrieval, browsing, etc. In this paper, we formulate video summarization as a sequence labeling problem. Unlike existing approaches that use recurrent models, we propose fully convolutional sequence models to solve video summarization. We firstly establish a novel connection between semantic segmentation and video summarization, and then adapt popular semantic segmentation networks for video summarization. Extensive experiments and analysis on two benchmark datasets demonstrate the effectiveness of our models.
Code (0)
등록된 구현이 없습니다.
Tasks
RetrievalSegmentationSemantic SegmentationVideo SummarizationSimilar Papers 제목 키워드 기반
Your Interest, Your Summaries: Query-Focused Long Video Summarization
Generating a concise and informative video summary from a long video is important, yet subjective due to varying scene importance. Users' ability to specify scene importance through text queries enhances the relevance of…
Query focused video summarizationVideo SummarizationReconstructive Sequence-Graph Network for Video Summarization
Exploiting the inner-shot and inter-shot dependencies is essential for key-shot based video summarization. Current approaches mainly devote to modeling the video as a frame sequence by recurrent neural networks. However,…
Video SummarizationA Novel Approach for Robust Multi Human Action Recognition and Summarization based on 3D Convolutional Neural Networks
Human actions in videos are 3D signals. However, there are a few methods available for multiple human action recognition. For long videos, it's difficult to search within a video for a specific action and/or person. For …
Action DetectionAction RecognitionTemporal Action LocalizationVideo SummarizationFullTransNet: Full Transformer with Local-Global Attention for Video Summarization
Video summarization mainly aims to produce a compact, short, informative, and representative synopsis of raw videos, which is of great importance for browsing, analyzing, and understanding video content. Dominant video s…
DecoderSupervised Video SummarizationVideo SummarizationUnsupervised Video Summarization with a Convolutional Attentive Adversarial Network
With the explosive growth of video data, video summarization, which attempts to seek the minimum subset of frames while still conveying the main story, has become one of the hottest topics. Nowadays, substantial achievem…
Generative Adversarial NetworkUnsupervised Video SummarizationVideo Summarization