paper-with-me

홈 › Papers

Learned Region Sparsity and Diversity Also Predicts Visual Attention

2016-12-01 · NeurIPS 2016 12 · Zijun Wei, Hossein Adeli, Minh Hoai Nguyen, Greg Zelinsky, Dimitris Samaras

Learned region sparsity has achieved state-of-the-art performance in classification tasks by exploiting and integrating a sparse set of local information into global decisions. The underlying mechanism resembles how people sample information from an image with their eye movements when making similar decisions. In this paper we incorporate the biologically plausible mechanism of Inhibition of Return into the learned region sparsity model, thereby imposing diversity on the selected regions. We investigate how these mechanisms of sparsity and diversity relate to visual attention by testing our model on three different types of visual search tasks. We report state-of-the-art results in predicting the locations of human gaze fixations, even though our model is trained only on image-level labels without object location annotations. Notably, the classification performance of the extended model remains the same as the original. This work suggests a new computational perspective on visual attention mechanisms and shows how the inclusion of attention-based mechanisms can improve computer vision techniques.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityGeneral Classification

Similar Papers 제목 키워드 기반

The Loupe: A Plug-and-Play Attention Module for Amplifying Discriminative Features in Vision Transformers

2025-08-20 · Naren Sengodan arxiv

Fine-Grained Visual Classification (FGVC) requires models to focus on subtle, task-relevant regions rather than broad object context. We present The Loupe, a lightweight plug-and-play spatial gating module for hierarchic…

4D-Former: Multimodal 4D Panoptic Segmentation

2023-11-02 · Ali Athar, Enxu Li, Sergio Casas, Raquel Urtasun

4D panoptic segmentation is a challenging but practically useful task that requires every point in a LiDAR point-cloud sequence to be assigned a semantic class label, and individual objects to be segmented and tracked ov…

4D Panoptic SegmentationPanoptic SegmentationPanoptic TrackingSegmentation

Deep Learning for Asynchronous Massive Access with Data Frame Length Diversity

2023-05-12 · Yanna Bai, Wei Chen, Bo Ai, Petar Popovski

Grant-free non-orthogonal multiple access has been regarded as a viable approach to accommodate access for a massive number of machine-type devices with small data packets. The sporadic activation of the devices creates …

Action DetectionActivity Detectioncompressed sensingDeep Learning+1

AINet: Anchor Instances Learning for Regional Heterogeneity in Whole Slide Image

2026-02-21 · Tingting Zheng, Hongxun Yao, Kui Jiang, Sicheng Zhao 외 arxiv

Recent advances in multi-instance learning (MIL) have witnessed impressive performance in whole slide image (WSI) analysis. However, the inherent sparsity of tumors and their morphological diversity lead to obvious heter…

Learned Contextual Feature Reweighting for Image Geo-Localization

2017-07-01 · CVPR 2017 7 · Hyo Jin Kim, Enrique Dunn, Jan-Michael Frahm

We address the problem of large scale image geo-localization where the location of an image is estimated by identifying geo-tagged reference images depicting the same place. We propose a novel model for learning image re…

geo-localization