paper-with-me

Papers

Graph-Theoretic Spatiotemporal Context Modeling for Video Saliency Detection

2017-07-25 · Lina Wei, Fangfang Wang, Xi Li, Fei Wu, Jun Xiao

As an important and challenging problem in computer vision, video saliency detection is typically cast as a spatiotemporal context modeling problem over consecutive frames. As a result, a key issue in video saliency detection is how to effectively capture the intrinsical properties of atomic video structures as well as their associated contextual interactions along the spatial and temporal dimensions. Motivated by this observation, we propose a graph-theoretic video saliency detection approach based on adaptive video structure discovery, which is carried out within a spatiotemporal atomic graph. Through graph-based manifold propagation, the proposed approach is capable of effectively modeling the semantically contextual interactions among atomic video structures for saliency detection while preserving spatial smoothness and temporal consistency. Experiments demonstrate the effectiveness of the proposed approach over several benchmark datasets.

📄 PDF Abstract BibTeX arXiv:1707.07815

Code (0)

등록된 구현이 없습니다.

Tasks

Saliency DetectionVideo Saliency Detection

Similar Papers 제목 키워드 기반

Language-guided Recursive Spatiotemporal Graph Modeling for Video Summarization

2025-09-06 · Jungin Park, Jiyoung Lee, Kwanghoon Sohn arxiv

Video summarization aims to select keyframes that are visually diverse and can represent the whole story of a given video. Previous approaches have focused on global interlinkability between frames in a video by temporal…

Video Summarization

High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation

2025-10-13 · Runyang Feng, Hyung Jin Chang, Tze Ho Elden Tse, Boeun Kim 외 arxiv

Modeling high-resolution spatiotemporal representations, including both global dynamic contexts (e.g., holistic human motion tendencies) and local motion details (e.g., high-frequency changes of keypoints), is essential …

Pose Estimation

Learning World Models for Interactive Video Generation

2025-05-28 · Taiye Chen, Xun Hu, Zihan Ding, Chi Jin

Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. However, present models for long video generation have limited inherent world mo…

In-Context LearningRetrievalRetrieval-augmented GenerationVideo Generation+1

Two-stream Spatiotemporal Feature for Video QA Task

2019-07-11 · Chiwan Song, Woobin Im, Sung-Eui Yoon

Understanding the content of videos is one of the core techniques for developing various helpful applications in the real world, such as recognizing various human actions for surveillance systems or customer behavior ana…

Action RecognitionTemporal Action LocalizationVocal Bursts Valence Prediction

Pruning for Generalization: A Transfer-Oriented Spatiotemporal Graph Framework

2026-02-04 · Zihao Jing, Yuxi Long, Ganlin Feng arxiv

Multivariate time series forecasting in graph-structured domains is critical for real-world applications, yet existing spatiotemporal models often suffer from performance degradation under data scarcity and cross-domain …

Multivariate Time Series Forecasting