paper-with-me

Papers

Visual Attention driven by Convolutional Features

2018-07-12 · Zanca Dario, Gori Marco

The understanding of where humans look in a scene is a problem of great interest in visual perception and computer vision. When eye-tracking devices are not a viable option, models of human attention can be used to predict fixations. In this paper we give two contribution. First, we show a model of visual attention that is simply based on deep convolutional neural networks trained for object classification tasks. A method for visualizing saliency maps is defined which is evaluated in a saliency prediction task. Second, we integrate the information of these maps with a bottom-up differential model of eye-movements to simulate visual attention scanpaths. Results on saliency prediction and scores of similarity with human scanpaths demonstrate the effectiveness of this model.

📄 PDF Abstract BibTeX arXiv:1807.10576

Code (0)

등록된 구현이 없습니다.

Tasks

Saliency Prediction

Similar Papers 제목 키워드 기반

Task-driven Webpage Saliency

2018-09-01 · ECCV 2018 9 · Quanlong Zheng, Jianbo Jiao, Ying Cao, Rynson W. H. Lau

In this paper, we present an end-to-end learning framework for predicting task-driven visual saliency on webpages. Given a webpage, we propose a convolutional neural network to predict where people look at it under diffe…

PredictionSaliency DetectionSaliency Prediction

A dynamic vision sensor object recognition model based on trainable event-driven convolution and spiking attention mechanism

2024-09-19 · Peng Zheng, Qian Zhou

Spiking Neural Networks (SNNs) are well-suited for processing event streams from Dynamic Visual Sensors (DVSs) due to their use of sparse spike-based coding and asynchronous event-driven computation. To extract features …

Object Recognition

An attention-driven hierarchical multi-scale representation for visual recognition

2021-10-23 · Zachary Wharton, Ardhendu Behera, Asish Bera

Convolutional Neural Networks (CNNs) have revolutionized the understanding of visual content. This is mainly due to their ability to break down an image into smaller pieces, extract multi-scale localized features and com…

Decision MakingFine-Grained Visual Recognition

STNet: Selective Tuning of Convolutional Networks for Object Localization

2017-08-21 · Mahdi Biparva, John Tsotsos

Visual attention modeling has recently gained momentum in developing visual hierarchies provided by Convolutional Neural Networks. Despite recent successes of feedforward processing on the abstraction of concepts form ra…

ObjectObject Localization

Landmark Guided Visual Feature Extractor for Visual Speech Recognition with Limited Resource

2025-08-10 · Lei Yang, Junshan Jin, Mingyuan Zhang, Yi He 외 arxiv

Visual speech recognition is a technique to identify spoken content in silent speech videos, which has raised significant attention in recent years. Advancements in data-driven deep learning methods have significantly im…

Visual Speech Recognition