paper-with-me

홈 › Papers

ON NEURAL NETWORK GENERALIZATION VIA PROMOTING WITHIN-LAYER ACTIVATION DIVERSITY

2021-01-01 · Firas Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef Gabbouj

During the last decade, neural networks have been intensively used to tackle various problems and they have often led to state-of-the-art results. These networks are composed of multiple jointly optimized layers arranged in a hierarchical structure. At each layer, the aim is to learn to extract hidden patterns needed to solve the problem at hand and forward it to the next layers. In the standard form, a neural network is trained with gradient-based optimization, where the errors are back-propagated from the last layer back to the first one. Thus at each optimization step, neurons at a given layer receive feedback from neurons belonging to higher layers of the hierarchy. In this paper, we propose to complement this traditional 'between-layer' feedback with additional 'within-layer' feedback to encourage diversity of the activations within the same layer. To this end, we measure the pairwise similarity between the outputs of the neurons and use it to model the layer's overall diversity. By penalizing similarities and promoting diversity, we encourage each neuron to learn a distinctive representation and, thus, to enrich the data representation learned within the layer and to increase the total capacity of the model. We theoretically study how the within-layer activation diversity affects the generalization performance of a neural network in a supervised context and we prove that increasing the diversity of hidden activations reduces the estimation error. In addition to the theoretical guarantees, we present an empirical study confirming that the proposed approach enhances the performance of neural networks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Improving Neural Network Generalization via Promoting Within-Layer Diversity

2021-09-29 · Firas Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef Gabbouj

Neural networks are composed of multiple layers arranged in a hierarchical structure jointly trained with a gradient-based optimization, where the errors are back-propagated from the last layer back to the first one. At …

Diversity

WLD-Reg: A Data-dependent Within-layer Diversity Regularizer

2023-01-03 · Firas Laakom, Jenni Raitoharju, Alexandros Iosifidis, Moncef Gabbouj

Neural networks are composed of multiple layers arranged in a hierarchical structure jointly trained with a gradient-based optimization, where the errors are back-propagated from the last layer back to the first one. At …

Diversity

Feature diversity in self-supervised learning

2022-09-02 · Pranshu Malviya, Arjun Vaithilingam Sudhakar

Many studies on scaling laws consider basic factors such as model size, model shape, dataset size, and compute power. These factors are easily tunable and represent the fundamental elements of any machine learning setup.…

DiversitySelf-Supervised Learning

Learning Latent Space Models with Angular Constraints

2017-08-01 · ICML 2017 8 · Pengtao Xie, Yuntian Deng, Yi Zhou, Abhimanu Kumar 외

The large model capacity of latent space models (LSMs) enables them to achieve great performance on various applications, but meanwhile renders LSMs to be prone to overfitting. Several recent studies investigate a n…

Diversity

Training Deep Fourier Neural Networks To Fit Time-Series Data

2014-05-09 · Michael S. Gashler, Stephen C. Ashmore

We present a method for training a deep neural network containing sinusoidal activation functions to fit to time-series data. Weights are initialized using a fast Fourier transform, then trained with regularization to im…

Time SeriesTime Series Analysis