paper-with-me

홈 › Papers

Understanding spatial correlation in eye-fixation maps for visual attention in videos

2019-01-30 · Tariq Alshawi, Zhiling Long, Ghassan AlRegib

In this paper, we present an analysis of recorded eye-fixation data from human subjects viewing video sequences. The purpose is to better understand visual attention for videos. Utilizing the eye-fixation data provided in the CRCNS (Collaborative Research in Computational Neuroscience) dataset, this paper focuses on the relation between the saliency of a pixel and that of its direct neighbors, without making any assumption about the structure of the eye-fixation maps. By employing some basic concepts from information theory, the analysis shows substantial correlation between the saliency of a pixel and the saliency of its neighborhood. The analysis also provides insights into the structure and dynamics of the eye-fixation maps, which can be very useful in understanding video saliency and its applications.

📄 PDF Abstract BibTeX arXiv:1901.10957

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks

2021-08-10 · Chetan Ralekar, Shubham Choudhary, Tapan Kumar Gandhi, Santanu Chaudhury

Human observers engage in selective information uptake when classifying visual patterns. The same is true of deep neural networks, which currently constitute the best performing artificial vision systems. Our goal is to …

Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps

2025-05-19 · Ziqi Wen, Jonathan Skaza, Shravan Murlidaran, William Y. Wang 외

Although models exist that predict human response times (RTs) in tasks such as target search and visual discrimination, the development of image-computable predictors for scene understanding time remains an open challeng…

Scene Understanding

Attention Alignment Between Humans and Vision-Language Models

2026-06-16 · Isaac R. Christian, Udith Haputhanthrige, Hanna Hornfeld, Declan Campbell 외 arxiv

Visual perception depends on top-down goals and bottom-up sensory mechanisms. Vision-language models implement both, allowing us to treat each component as a separable hypothesis about what drives where we look. We compa…

Salient Object Detection Driven by Fixation Prediction

2018-06-01 · CVPR 2018 6 · Wenguan Wang, Jianbing Shen, Xingping Dong, Ali Borji

Research in visual saliency has been focused on two major types of models namely fixation prediction and salient object detection. The relationship between the two, however, has been less explored. In this paper, we prop…

Objectobject-detectionObject DetectionPrediction+3

Line Drawings of Natural Scenes Guide Visual Attention

2019-12-19 · Kai-Fu Yang, Wen-Wen Jiang, Teng-Fei Zhan, Yong-Jie Li

Visual search is an important strategy of the human visual system for fast scene perception. The guided search theory suggests that the global layout or other top-down sources of scenes play a crucial role in guiding obj…