paper-with-me

홈 › Papers

GAttANet: Global attention agreement for convolutional neural networks

2021-04-12 · Rufin VanRullen, Andrea Alamia

Transformer attention architectures, similar to those developed for natural language processing, have recently proved efficient also in vision, either in conjunction with or as a replacement for convolutional layers. Typically, visual attention is inserted in the network architecture as a (series of) feedforward self-attention module(s), with mutual key-query agreement as the main selection and routing operation. However efficient, this strategy is only vaguely compatible with the way that attention is implemented in biological brains: as a separate and unified network of attentional selection regions, receiving inputs from and exerting modulatory influence on the entire hierarchy of visual regions. Here, we report experiments with a simple such attention system that can improve the performance of standard convolutional networks, with relatively few additional parameters. Each spatial position in each layer of the network produces a key-query vector pair; all queries are then pooled into a global attention query. On the next iteration, the match between each key and the global attention query modulates the network's activations -- emphasizing or silencing the locations that agree or disagree (respectively) with the global attention system. We demonstrate the usefulness of this brain-inspired Global Attention Agreement network (GAttANet) for various convolutional backbones (from a simple 5-layer toy model to a standard ResNet50 architecture) and datasets (CIFAR10, CIFAR100, Imagenet-1k). Each time, our global attention system improves accuracy over the corresponding baseline.

📄 PDF Abstract BibTeX arXiv:2104.05575

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GASPnet: Global Agreement to Synchronize Phases

2025-07-22 · Andrea Alamia, Sabine Muzellec, Thomas Serre, Rufin VanRullen arxiv

In recent years, Transformer architectures have revolutionized most fields of artificial intelligence, relying on an attentional mechanism based on the agreement between keys and queries to select and route information i…

Near Real-Time Dust Aerosol Detection with 3D Convolutional Neural Networks on MODIS Data

2025-09-07 · Caleb Gates, Patrick Moorhead, Jayden Ferguson, Omar Darwish 외 arxiv

Dust storms harm health and reduce visibility; quick detection from satellites is needed. We present a near real-time system that flags dust at the pixel level using multi-band images from NASA's Terra and Aqua (MODIS). …

Why Self-Attention? A Targeted Evaluation of Neural Machine Translation Architectures

2018-08-27 · EMNLP 2018 10 · Gongbo Tang, Mathias Müller, Annette Rios, Rico Sennrich

Recently, non-recurrent architectures (convolutional, self-attentional) have outperformed RNNs in neural machine translation. CNNs and self-attentional networks can connect distant words via shorter network paths than RN…

Machine TranslationTranslationWord Sense Disambiguation

Conditionally Learn to Pay Attention for Sequential Visual Task

2019-11-11 · Jun He, Quan-Jie Cao, Lei Zhang

Sequential visual task usually requires to pay attention to its current interested object conditional on its previous observations. Different from popular soft attention mechanism, we propose a new attention framework by…

Language ModelingLanguage Modelling

HDAM: Heuristic Difference Attention Module for Convolutional Neural Networks

2022-02-19 · Yu Xue, Ziming Yuan

The attention mechanism is one of the most important priori knowledge to enhance convolutional neural networks. Most attention mechanisms are bound to the convolutional layer and use local or global contextual informatio…