paper-with-me

홈 › Papers

Deep Learning for Saliency Prediction in Natural Video

2016-04-27 · Souad Chaabouni, Jenny Benois-Pineau, Ofer Hadar, Chokri Ben Amar

The purpose of this paper is the detection of salient areas in natural video by using the new deep learning techniques. Salient patches in video frames are predicted first. Then the predicted visual fixation maps are built upon them. We design the deep architecture on the basis of CaffeNet implemented with Caffe toolkit. We show that changing the way of data selection for optimisation of network parameters, we can save computation cost up to 12 times. We extend deep learning approaches for saliency prediction in still images with RGB values to specificity of video using the sensitivity of the human visual system to residual motion. Furthermore, we complete primary colour pixel values by contrast features proposed in classical visual attention prediction models. The experiments are conducted on two publicly available datasets. The first is IRCCYN video database containing 31 videos with an overall amount of 7300 frames and eye fixations of 37 subjects. The second one is HOLLYWOOD2 provided 2517 movie clips with the eye fixations of 19 subjects. On IRCYYN dataset, the accuracy obtained is of 89.51%. On HOLLYWOOD2 dataset, results in prediction of saliency of patches show the improvement up to 2% with regard to RGB use only. The resulting accuracy of 76, 6% is obtained. The AUC metric in comparison of predicted saliency maps with visual fixation maps shows the increase up to 16% on a sample of video clips from this dataset.

📄 PDF Abstract BibTeX arXiv:1604.08010

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningPredictionSaliency PredictionSpecificity

Similar Papers 제목 키워드 기반

Audiovisual Saliency Prediction in Uncategorized Video Sequences based on Audio-Video Correlation

2021-01-07 · Maryam Qamar Butt, Anis Ur Rahman

Substantial research has been done in saliency modeling to develop intelligent machines that can perceive and interpret their surroundings. But existing models treat videos as merely image sequences excluding any audio i…

Saliency Prediction

A Learning-Based Visual Saliency Prediction Model for Stereoscopic 3D Video (LBVS-3D)

2018-03-13 · Amin Banitalebi-Dehkordi, Mahsa T. Pourazad, Panos Nasiopoulos

Over the past decade, many computational saliency prediction models have been proposed for 2D images and videos. Considering that the human visual system has evolved in a natural 3D environment, it is only natural to wan…

Saliency PredictionVideo Saliency Prediction

Viewport Prediction for Volumetric Video Streaming by Exploring Video Saliency and Trajectory Information

2023-11-28 · Jie Li, Zhixin Li, Zhi Liu, Pengyuan Zhou 외

Volumetric video, also known as hologram video, is a novel medium that portrays natural content in Virtual Reality (VR), Augmented Reality (AR), and Mixed Reality (MR). It is expected to be the next-gen video technology …

Mixed RealityPredictionSaliency Detection

Learning to Predict Salient Faces: A Novel Visual-Audio Saliency Model

2021-03-29 · ECCV 2020 8 · Yufan Liu, Minglang Qiao, Mai Xu, Bing Li 외

Recently, video streams have occupied a large proportion of Internet traffic, most of which contain human faces. Hence, it is necessary to predict saliency on multiple-face videos, which can provide attention cues for ma…

Saliency Prediction

Model-guided Multi-path Knowledge Aggregation for Aerial Saliency Prediction

2018-11-14 · Kui Fu, Jia Li, Yu Zhang, Hongze Shen 외

As an emerging vision platform, a drone can look from many abnormal viewpoints which brings many new challenges into the classic vision task of video saliency prediction. To investigate these challenges, this paper propo…

Aerial Video Saliency PredictionPredictionSaliency PredictionTransfer Learning+1