paper-with-me

Papers

A Learning-Based Visual Saliency Prediction Model for Stereoscopic 3D Video (LBVS-3D)

2018-03-13 · Amin Banitalebi-Dehkordi, Mahsa T. Pourazad, Panos Nasiopoulos

Over the past decade, many computational saliency prediction models have been proposed for 2D images and videos. Considering that the human visual system has evolved in a natural 3D environment, it is only natural to want to design visual attention models for 3D content. Existing monocular saliency models are not able to accurately predict the attentive regions when applied to 3D image/video content, as they do not incorporate depth information. This paper explores stereoscopic video saliency prediction by exploiting both low-level attributes such as brightness, color, texture, orientation, motion, and depth, as well as high-level cues such as face, person, vehicle, animal, text, and horizon. Our model starts with a rough segmentation and quantifies several intuitive observations such as the effects of visual discomfort level, depth abruptness, motion acceleration, elements of surprise, size and compactness of the salient regions, and emphasizing only a few salient objects in a scene. A new fovea-based model of spatial distance between the image regions is adopted for considering local and global feature calculations. To efficiently fuse the conspicuity maps generated by our method to one single saliency map that is highly correlated with the eye-fixation data, a random forest based algorithm is utilized. The performance of the proposed saliency model is evaluated against the results of an eye-tracking experiment, which involved 24 subjects and an in-house database of 61 captured stereoscopic videos. Our stereo video database as well as the eye-tracking data are publicly available along with this paper. Experiment results show that the proposed saliency prediction method achieves competitive performance compared to the state-of-the-art approaches.

📄 PDF Abstract BibTeX arXiv:1803.04842

Code (1)

https://gitlab.com/abanitalebi/lbvs-3d 공식 구현

Tasks

Saliency PredictionVideo Saliency Prediction

Similar Papers 제목 키워드 기반

Learning to Explore Intrinsic Saliency for Stereoscopic Video

2019-06-01 · CVPR 2019 6 · Qiudan Zhang, Xu Wang, Shiqi Wang, Shikai Li 외

The human visual system excels at biasing the stereoscopic visual signals by the attention mechanisms. Traditional methods relying on the low-level features and depth relevant information for stereoscopic video saliency …

Saliency DetectionSaliency PredictionVideo Saliency DetectionVideo Saliency Prediction

Benchmark 3D eye-tracking dataset for visual saliency prediction on stereoscopic 3D video

2018-03-13

Visual Attention Models (VAMs) predict the location of an image or video regions that are most likely to attract human attention. Although saliency detection is well explored for 2D image and video content, there are onl…

Saliency DetectionSaliency PredictionVideo Saliency Prediction

A Learning-Based Visual Saliency Fusion Model for High Dynamic Range Video (LBVS-HDR)

2018-03-13 · Amin Banitalebi-Dehkordi, Yuanyuan Dong, Mahsa T. Pourazad, Panos Nasiopoulos

Saliency prediction for Standard Dynamic Range (SDR) videos has been well explored in the last decade. However, limited studies are available on High Dynamic Range (HDR) Visual Attention Models (VAMs). Considering that t…

Saliency Prediction

Saliency Inspired Quality Assessment of Stereoscopic 3D Video

2018-03-12

To study the visual attentional behavior of Human Visual System (HVS) on 3D content, eye tracking experiments are performed and Visual Attention Models (VAMs) are designed. One of the main applications of these VAMs is i…

How do people explore virtual environments?

2016-12-13 · Vincent Sitzmann, Ana Serrano, Amy Pavel, Maneesh Agrawala 외

Understanding how people explore immersive virtual environments is crucial for many applications, such as designing virtual reality (VR) content, developing new compression algorithms, or learning computational models of…

Video Synopsis