paper-with-me

홈 › Papers

TVSum: Summarizing Web Videos Using Titles

2015-06-01 · CVPR 2015 6 · Yale Song, Jordi Vallmitjana, Amanda Stent, Alejandro Jaimes

Video summarization is a challenging problem in part because knowing which part of a video is important requires prior knowledge about its main topic. We present TVSum, an unsupervised video summarization framework that uses title-based image search results to find visually important shots. We observe that a video title is often carefully chosen to be maximally descriptive of its main topic, and hence images related to the title can serve as a proxy for important visual concepts of the main topic. However, because titles are free-formed, unconstrained, and often written ambiguously, images searched using the title can contain noise (images irrelevant to video content) and variance (images of different topics). To deal with this challenge, we developed a novel co-archetypal analysis technique that learns canonical visual concepts shared between video and images, but not in either alone, by finding a joint-factorial representation of two data sets. We introduce a new benchmark dataset, TVSum50, that contains 50 videos and their shot-level importance scores annotated via crowdsourcing. Experimental results on two datasets, SumMe and TVSum50, suggest our approach produces superior quality summaries compared to several recently proposed approaches.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DescriptiveImage RetrievalUnsupervised Video SummarizationVideo Summarization

Similar Papers 제목 키워드 기반

Summarizing Videos with Attention

2018-12-05 · Jiri Fajtl, Hajar Sadeghi Sokeh, Vasileios Argyriou, Dorothy Monekosso 외

In this work we propose a novel method for supervised, keyshots based video summarization by applying a conceptually simple and computationally efficient soft, self-attention mechanism. Current state of the art methods l…

Video Summarization

Does SpatioTemporal information benefit Two video summarization benchmarks?

2024-10-04 · Aashutosh Ganesh, Mirela Popa, Daan Odijk, Nava Tintarev

An important aspect of summarizing videos is understanding the temporal context behind each part of the video to grasp what is and is not important. Video summarization models have in recent years modeled spatio-temporal…

Activity RecognitionVideo Summarization

Key Frame Extraction with Attention Based Deep Neural Networks

2023-06-21 · Samed Arslan, Senem Tanberk

Automatic keyframe detection from videos is an exercise in selecting scenes that can best summarize the content for long videos. Providing a summary of the video is an important task to facilitate quick browsing and cont…

Video RetrievalVideo Summarization

Demo-ICL: In-Context Learning for Procedural Video Knowledge Acquisition

2026-02-09 · Yuhao Dong, Shulin Tian, Shuai Liu, Shuangrui Ding 외 arxiv

Despite the growing video understanding capabilities of recent Multimodal Large Language Models (MLLMs), existing video benchmarks primarily assess understanding based on models' static, internal knowledge, rather than t…

Role of Audio in Audio-Visual Video Summarization

2022-12-02 · Ibrahim Shoer, Berkay Kopru, Engin Erzin

Video summarization attracts attention for efficient video representation, retrieval, and browsing to ease volume and traffic surge problems. Although video summarization mostly uses the visual channel for compaction, th…

RetrievalVideo Summarization