paper-with-me

홈 › Papers

Video Skimming: Taxonomy and Comprehensive Survey

2019-09-21 · Vivekraj V. K., Debashis Sen, Balasubramanian Raman

Video skimming, also known as dynamic video summarization, generates a temporally abridged version of a given video. Skimming can be achieved by identifying significant components either in uni-modal or multi-modal features extracted from the video. Being dynamic in nature, video skimming, through temporal connectivity, allows better understanding of the video from its summary. Having this obvious advantage, recently, video skimming has drawn the focus of many researchers benefiting from the easy availability of the required computing resources. In this paper, we provide a comprehensive survey on video skimming focusing on the substantial amount of literature from the past decade. We present a taxonomy of video skimming approaches, and discuss their evolution highlighting key advances. We also provide a study on the components required for the evaluation of a video skimming performance.

📄 PDF Abstract BibTeX arXiv:1909.12948

Code (0)

등록된 구현이 없습니다.

Tasks

SurveyVideo Summarization

Similar Papers 제목 키워드 기반

A Survey of Single-Scene Video Anomaly Detection

2020-04-13 · Bharathkumar Ramachandra, Michael J. Jones, Ranga Raju Vatsavai

This survey article summarizes research trends on the topic of anomaly detection in video feeds of a single scene. We discuss the various problem formulations, publicly available datasets and evaluation criteria. We cate…

Anomaly DetectionSurveyVideo Anomaly Detection

A Comprehensive Review of Few-shot Action Recognition

2024-07-20 · Yuyang Wanyan, Xiaoshan Yang, WeiMing Dong, Changsheng Xu

Few-shot action recognition aims to address the high cost and impracticality of manually labeling complex and variable video data in action recognition. It requires accurately classifying human actions in videos using on…

Action RecognitionFew-Shot action recognitionFew Shot Action RecognitionFew-Shot Learning+6

Distorted or Fabricated? A Survey on Hallucination in Video LLMs

2026-04-14 · Yiyang Huang, Yitian Zhang, Yizhou Wang, Mingyuan Zhang 외 arxiv

Despite significant progress in video-language modeling, hallucinations remain a persistent challenge in Video Large Language Models (Vid-LLMs), referring to outputs that appear plausible yet contradict the content of th…

Visual Grounding

A Survey on Future Frame Synthesis: Bridging Deterministic and Generative Approaches

2024-01-26 · Ruibo Ming, Zhewei Huang, Jingwei Wu, Zhuoxuan Ju 외

Future Frame Synthesis (FFS, aka Video Frame Prediction) focuses on generating future frame sequences conditioned on existing content. This survey provides a comprehensive review of existing research on FFS, covering com…

SurveyVideo Prediction

Vision Mamba: A Comprehensive Survey and Taxonomy

2024-05-07 · Xiao Liu, Chenxu Zhang, Lei Zhang

State Space Model (SSM) is a mathematical model used to describe and analyze the behavior of dynamic systems. This model has witnessed numerous applications in several fields, including control theory, signal processing,…

MambaMedical Image AnalysisState Space ModelsSurvey+2