On Translation Invariance in CNNs: Convolutional Layers can Exploit Absolute Spatial Location
In this paper we challenge the common assumption that convolutional layers in modern CNNs are translation invariant. We show that CNNs can and will exploit the absolute spatial location by learning filters that respond exclusively to particular absolute locations by exploiting image boundary effects. Because modern CNNs filters have a huge receptive field, these boundary effects operate even far from the image boundary, allowing the network to exploit absolute spatial location all over the image. We give a simple solution to remove spatial location encoding which improves translation invariance and thus gives a stronger visual inductive bias which particularly benefits small data sets. We broadly demonstrate these benefits on several architectures and various applications such as image classification, patch matching, and two video classification datasets.
Code (3)
Tasks
General Classificationimage-classificationImage ClassificationInductive BiasPatch MatchingSmall Data Image ClassificationTranslationVideo ClassificationSimilar Papers 제목 키워드 기반
Tracking translation invariance in CNNs
Although Convolutional Neural Networks (CNNs) are widely used, their translation invariance (ability to deal with translated inputs) is still subject to some controversy. We explore this question using translation-sensit…
SensitivityTranslationQuantifying Translation-Invariance in Convolutional Neural Networks
A fundamental problem in object recognition is the development of image representations that are invariant to common transformations such as translation, rotation, and small deformations. There are multiple hypotheses re…
Data AugmentationObject RecognitionTranslationSpectral Networks and Locally Connected Networks on Graphs
Convolutional Neural Networks are extremely efficient architectures in image and audio recognition tasks, thanks to their ability to exploit the local translational invariance of signal classes over their domain. In this…
ClusteringTranslationRevisiting Data Augmentation for Rotational Invariance in Convolutional Neural Networks
Convolutional Neural Networks (CNN) offer state of the art performance in various computer vision tasks. Many of those tasks require different subtypes of affine invariances (scale, rotational, translational) to image tr…
Data Augmentationimage-classificationImage ClassificationDeep Learning for Target Classification from SAR Imagery: Data Augmentation and Translation Invariance
This report deals with translation invariance of convolutional neural networks (CNNs) for automatic target recognition (ATR) from synthetic aperture radar (SAR) imagery. In particular, the translation invariance of CNNs …
Data AugmentationGeneral ClassificationTranslation