paper-with-me

홈 › Papers

A ReLU Dense Layer to Improve the Performance of Neural Networks

2020-10-22 · Alireza M. Javid, Sandipan Das, Mikael Skoglund, Saikat Chatterjee

We propose ReDense as a simple and low complexity way to improve the performance of trained neural networks. We use a combination of random weights and rectified linear unit (ReLU) activation function to add a ReLU dense (ReDense) layer to the trained neural network such that it can achieve a lower training loss. The lossless flow property (LFP) of ReLU is the key to achieve the lower training loss while keeping the generalization error small. ReDense does not suffer from vanishing gradient problem in the training due to having a shallow structure. We experimentally show that ReDense can improve the training and testing performance of various neural network architectures with different optimization loss and activation functions. Finally, we test ReDense on some of the state-of-the-art architectures and show the performance improvement on benchmark datasets.

📄 PDF Abstract BibTeX arXiv:2010.13572

Code (1)

alirezamj14/ReDense 공식 구현 tf

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…

Similar Papers 제목 키워드 기반

Phase diagram for two-layer ReLU neural networks at infinite-width limit

2020-07-15 · Tao Luo, Zhi-Qin John Xu, Zheng Ma, Yaoyu Zhang

How neural network behaves during the training over different choices of hyperparameters is an important question in the study of neural networks. In this work, inspired by the phase diagram in statistical mechanics, we …

Empirical Phase Diagram for Three-layer Neural Networks with Infinite Width

2022-05-24 · Hanxu Zhou, Qixuan Zhou, Zhenyuan Jin, Tao Luo 외

Substantial work indicates that the dynamics of neural networks (NNs) is closely related to their initialization of parameters. Inspired by the phase diagram for two-layer ReLU NNs with infinite width (Luo et al., 2021),…

EraseReLU: A Simple Way to Ease the Training of Deep Convolution Neural Networks

2017-09-22 · Xuanyi Dong, Guoliang Kang, Kun Zhan, Yi Yang

For most state-of-the-art architectures, Rectified Linear Unit (ReLU) becomes a standard component accompanied with each layer. Although ReLU can ease the network training to an extent, the character of blocking negative…

BlockingImage Classification

AReLU: Attention-based Rectified Linear Unit

2020-06-24 · Dengsheng Chen, Jun Li, Kai Xu

Element-wise activation functions play a critical role in deep neural networks via affecting the expressivity power and the learning dynamics. Learning-based activation functions have recently gained increasing attention…

Meta-LearningTransfer Learning

The Effect of Residual Architecture on the Per-Layer Gradient of Deep Networks

2019-09-25 · Etai Littwin, Lior Wolf

A critical part of the training process of neural networks takes place in the very first gradient steps post initialization. In this work, we study the connection between the network's architecture and initialization par…