paper-with-me

Papers

Predicting Salient Face in Multiple-Face Videos

2017-07-01 · CVPR 2017 7 · Yufan Liu, Songyang Zhang, Mai Xu, Xuming He

Although the recent success of convolutional neural network (CNN) advances state-of-the-art saliency prediction in static images, few work has addressed the problem of predicting attention in videos. On the other hand, we find that the attention of different subjects consistently focuses on a single face in each frame of videos involving multiple faces. Therefore, we propose in this paper a novel deep learning (DL) based method to predict salient face in multiple-face videos, which is capable of learning features and transition of salient faces across video frames. In particular, we first learn a CNN for each frame to locate salient face. Taking CNN features as input, we develop a multiple-stream long short-term memory (M-LSTM) network to predict the temporal transition of salient faces in video sequences. To evaluate our DL-based method, we build a new eye-tracking database of multiple-face videos. The experimental results show that our method outperforms the prior state-of-the-art methods in predicting visual attention on faces in multiple-face videos.

📄 PDF Abstract BibTeX

Code (1)

yufanLIU/salient-face-in-MUVFET 공식 구현 tf

Tasks

Saliency Prediction

Similar Papers 제목 키워드 기반

Saliency-guided Emotion Modeling: Predicting Viewer Reactions from Video Stimuli

2025-05-25 · Akhila Yaragoppa, Siddharth

Understanding the emotional impact of videos is crucial for applications in content creation, advertising, and Human-Computer Interaction (HCI). Traditional affective computing methods rely on self-reported emotions, fac…

Face Video Deblurring Using 3D Facial Priors

2019-10-01 · ICCV 2019 10 · Wenqi Ren, Jiaolong Yang, Senyou Deng, David Wipf 외

Existing face deblurring methods only consider single frames and do not account for facial structure and identity information. These methods struggle to deblur face videos that exhibit significant pose variations and mis…

3D Face ReconstructionDeblurringDecoderFace Reconstruction+1

Learning to Predict Salient Faces: A Novel Visual-Audio Saliency Model

2021-03-29 · ECCV 2020 8 · Yufan Liu, Minglang Qiao, Mai Xu, Bing Li 외

Recently, video streams have occupied a large proportion of Internet traffic, most of which contain human faces. Hence, it is necessary to predict saliency on multiple-face videos, which can provide attention cues for ma…

Saliency Prediction

Multi-modality Deep Restoration of Extremely Compressed Face Videos

2021-07-05 · Xi Zhang, Xiaolin Wu

Arguably the most common and salient object in daily video communications is the talking head, as encountered in social media, virtual classrooms, teleconferences, news broadcasting, talk shows, etc. When communication b…

Quantization

LMME3DHF: Benchmarking and Evaluating Multimodal 3D Human Face Generation with LMMs

2025-04-29 · Woo Yi Yang, Jiarui Wang, Sijing Wu, Huiyu Duan 외

The rapid advancement in generative artificial intelligence have enabled the creation of 3D human faces (HFs) for applications including media production, virtual reality, security, healthcare, and game development, etc.…

BenchmarkingFace GenerationQuestion AnsweringSaliency Prediction+1