paper-with-me

Papers

Unified Dynamic Scanpath Predictors Outperform Individually Trained Neural Models

2024-05-05 · Fares Abawi, Di Fu, Stefan Wermter

Previous research on scanpath prediction has mainly focused on group models, disregarding the fact that the scanpaths and attentional behaviors of individuals are diverse. The disregard of these differences is especially detrimental to social human-robot interaction, whereby robots commonly emulate human gaze based on heuristics or predefined patterns. However, human gaze patterns are heterogeneous and varying behaviors can significantly affect the outcomes of such human-robot interactions. To fill this gap, we developed a deep learning-based social cue integration model for saliency prediction to instead predict scanpaths in videos. Our model learned scanpaths by recursively integrating fixation history and social cues through a gating mechanism and sequential attention. We evaluated our approach on gaze datasets of dynamic social scenes, observed under the free-viewing condition. The introduction of fixation history into our models makes it possible to train a single unified model rather than the resource-intensive approach of training individual models for each set of scanpaths. We observed that the late neural integration approach surpasses early fusion when training models on a large dataset, in comparison to a smaller dataset with a similar distribution. Results also indicate that a single unified model, trained on all the observers' scanpaths, performs on par or better than individually trained models. We hypothesize that this outcome is a result of the group saliency representations instilling universal attention in the model, while the supervisory signal and fixation history guide it to learn personalized attentional behaviors, providing the unified model a benefit over individual models due to its implicit representation of universal attention.

📄 PDF Abstract BibTeX arXiv:2405.02929

Code (0)

등록된 구현이 없습니다.

Tasks

Saliency PredictionScanpath prediction

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Scanpath Prediction on Information Visualisations

2021-12-04 · Yao Wang, Mihai Bâce, Andreas Bulling

We propose Unified Model of Saliency and Scanpaths (UMSS) -- a model that learns to predict visual saliency and scanpaths (i.e. sequences of eye fixations) on information visualisations. Although scanpaths provide rich i…

PredictionSaliency PredictionScanpath prediction

Influence of initial fixation position in scene viewing

2016-07-13

During scene perception our eyes generate complex sequences of fixations. Predictors of fixation locations are bottom-up factors like luminance contrast, top-down factors like viewing instruction, and systematic biases l…

Position

Contrastive Language-Image Pretrained Models are Zero-Shot Human Scanpath Predictors

2023-05-21 · Dario Zanca, Andrea Zugarini, Simon Dietz, Thomas R. Altstidl 외

Understanding the mechanisms underlying human attention is a fundamental challenge for both vision science and artificial intelligence. While numerous computational models of free-viewing have been proposed, less is know…

Decision MakingScanpath prediction

CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling

2025-07-16 · Trong-Thang Pham, Akash Awasthi, Saba Khan, Esteban Duran Marti 외 arxiv

Understanding radiologists' eye movement during Computed Tomography (CT) reading is crucial for developing effective interpretable computer-aided diagnosis systems. However, CT research in this area has been limited by t…

Scanpath prediction

A Probabilistic Time-Evolving Approach to Scanpath Prediction

2022-04-20 · Daniel Martin, Diego Gutierrez, Belen Masia

Human visual attention is a complex phenomenon that has been studied for decades. Within it, the particular problem of scanpath prediction poses a challenge, particularly due to the inter- and intra-observer variability,…

Dynamic Time WarpingPredictionScanpath prediction