paper-with-me

홈 › Papers

On the existence of optimal shallow feedforward networks with ReLU activation

2023-03-06 · Steffen Dereich, Sebastian Kassing

We prove existence of global minima in the loss landscape for the approximation of continuous target functions using shallow feedforward artificial neural networks with ReLU activation. This property is one of the fundamental artifacts separating ReLU from other commonly used activation functions. We propose a kind of closure of the search space so that in the extended space minimizers exist. In a second step, we show under mild assumptions that the newly added functions in the extension perform worse than appropriate representable ReLU networks. This then implies that the optimal response in the extended target space is indeed the response of a ReLU network.

📄 PDF Abstract BibTeX arXiv:2303.03950

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the existence of minimizers in shallow residual ReLU neural network optimization landscapes

2023-02-28 · Steffen Dereich, Arnulf Jentzen, Sebastian Kassing

In this article, we show existence of minimizers in the loss landscape for residual artificial neural networks (ANNs) with multi-dimensional input layer and one hidden layer with ReLU activation. Our work contrasts earli…

Math

Optimizing Performance of Feedforward and Convolutional Neural Networks through Dynamic Activation Functions

2023-08-10 · Chinmay Rane, Kanishka Tyagi, Michael Manry

Deep learning training training algorithms are a huge success in recent years in many fields including speech, text,image video etc. Deeper and deeper layers are proposed with huge success with resnet structures having a…

Tropical Geometry of Deep Neural Networks

2018-05-18 · ICML 2018 7 · Liwen Zhang, Gregory Naitzat, Lek-Heng Lim

We establish, for the first time, connections between feedforward neural networks with ReLU activation and tropical geometry --- we show that the family of such neural networks is equivalent to the family of tropical rat…

Generalized Activation via Multivariate Projection

2023-09-29 · Jiayun Li, Yuxiao Cheng, Yiwen Lu, Zhuofan Xia 외

Activation functions are essential to introduce nonlinearity into neural networks, with the Rectified Linear Unit (ReLU) often favored for its simplicity and effectiveness. Motivated by the structural similarity between …

Can Implicit Bias Imply Adversarial Robustness?

2024-05-24 · Hancheng Min, René Vidal

The implicit bias of gradient-based training algorithms has been considered mostly beneficial as it leads to trained networks that often generalize well. However, Frei et al. (2023) show that such implicit bias can harm …

Adversarial Robustness