Learning From Unique Perspectives: User-Aware Saliency Modeling
Everyone is unique. Given the same visual stimuli, people's attention is driven by both salient visual cues and their own inherent preferences. Knowledge of visual preferences not only facilitates understanding of fine-grained attention patterns of diverse users, but also has the potential of benefiting the development of customized applications. Nevertheless, existing saliency models typically limit their scope to attention as it applies to the general population and ignore the variability between users' behaviors. In this paper, we identify the critical roles of visual preferences in attention modeling, and for the first time study the problem of user-aware saliency modeling. Our work aims to advance attention research from three distinct perspectives: (1) We present a new model with the flexibility to capture attention patterns of various combinations of users, so that we can adaptively predict personalized attention, user group attention, and general saliency at the same time with one single model; (2) To augment models with knowledge about the composition of attention from different users, we further propose a principled learning method to understand visual attention in a progressive manner; and (3) We carry out extensive analyses on publicly available saliency datasets to shed light on the roles of visual preferences. Experimental results on diverse stimuli, including naturalistic images and web pages, demonstrate the advantages of our method in capturing the distinct visual behaviors of different users and the general saliency of visual stimuli.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Frequency-aware Graph Signal Processing for Collaborative Filtering
Graph Signal Processing (GSP) based recommendation algorithms have recently attracted lots of attention due to its high efficiency. However, these methods failed to consider the importance of various interactions that re…
Collaborative FilteringA Comparative Study on Textual Saliency of Styles from Eye Tracking, Annotations, and Language Models
There is growing interest in incorporating eye-tracking data and other implicit measures of human language processing into natural language processing (NLP) pipelines. The data from human language processing contain uniq…
Few-Shot LearningGenerating Diversified Comments via Reader-Aware Topic Modeling and Saliency Detection
Automatic comment generation is a special and challenging task to verify the model ability on news content comprehension and language generation. Comments not only convey salient and interesting information in news artic…
ArticlesClusteringComment GenerationDecoder+3SalFoM: Dynamic Saliency Prediction with Video Foundation Models
Recent advancements in video saliency prediction (VSP) have shown promising performance compared to the human visual system, whose emulation is the primary goal of VSP. However, current state-of-the-art models employ spa…
DecoderPredictionSaliency PredictionVideo Saliency PredictionCombining Visual Saliency Methods and Sparse Keypoint Annotations to Providently Detect Vehicles at Night
Provident detection of other road users at night has the potential for increasing road safety. For this purpose, humans intuitively use visual cues, such as light cones and light reflections emitted by other road users t…
object-detectionObject Detectionvehicle detection