paper-with-me

Papers

Learning Video Saliency from Human Gaze Using Candidate Selection

2013-06-01 · CVPR 2013 6 · Dmitry Rudoy, Dan B. Goldman, Eli Shechtman, Lihi Zelnik-Manor

During recent years remarkable progress has been made in visual saliency modeling. Our interest is in video saliency. Since videos are fundamentally different from still images, they are viewed differently by human observers. For example, the time each video frame is observed is a fraction of a second, while a still image can be viewed leisurely. Therefore, video saliency estimation methods should differ substantially from image saliency methods. In this paper we propose a novel method for video saliency estimation, which is inspired by the way people watch videos. We explicitly model the continuity of the video by predicting the saliency map of a given frame, conditioned on the map from the previous frame. Furthermore, accuracy and computation speed are improved by restricting the salient locations to a carefully selected candidate set. We validate our method using two gaze-tracked video datasets and show we outperform the state-of-the-art.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Saliency Prediction

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Measuring the Importance of Temporal Features in Video Saliency

2020-08-01 · ECCV 2020 8 · Matthias Tangemann, Matthias Kümmerer, Thomas S. A. Wallis, Matthias Bethge

Where people look when watching videos is believed to be heavily influenced by temporal patterns. In this work, we test this assumption by quantifying to which extent gaze on recent video saliency benchmarks can be predi…

Infinite Gaze Generation for Videos with Autoregressive Diffusion

2026-03-26 · Jenna Kang, Colin Groth, Tong Wu, Finley Torrens 외 arxiv

Predicting human gaze in video is fundamental to advancing scene understanding and multimodal interaction. While traditional saliency maps provide spatial probability distributions and scanpaths offer ordered fixations, …

Scene Understanding

Following Gaze in Video

2017-10-01 · ICCV 2017 10 · Adria Recasens, Carl Vondrick, Aditya Khosla, Antonio Torralba

Following the gaze of people inside videos is an important signal for understanding people and their actions. In this paper, we present an approach for following gaze in video by predicting where a person (in the video) …

Gaze Prediction in Dynamic 360° Immersive Videos

2018-06-01 · CVPR 2018 6 · Yanyu Xu, Yanbing Dong, Junru Wu, Zhengzhong Sun 외

This paper explores gaze prediction in dynamic $360^circ$ immersive videos, emph{i.e.}, based on the history scan path and VR contents, we predict where a viewer will look at an upcoming time. To tackle this problem, we …

Gaze PredictionPrediction

Improving saliency models' predictions of the next fixation with humans' intrinsic cost of gaze shifts

2022-07-09 · Florian Kadner, Tobias Thomas, David Hoppe, Constantin A. Rothkopf

The human prioritization of image regions can be modeled in a time invariant fashion with saliency maps or sequentially with scanpath models. However, while both types of models have steadily improved on several benchmar…

Decision MakingSequential Decision Making