paper-with-me

홈 › Papers

Improving the Expressive Power of Deep Neural Networks through Integral Activation Transform

2023-12-19 · Zezhong Zhang, Feng Bao, Guannan Zhang

The impressive expressive power of deep neural networks (DNNs) underlies their widespread applicability. However, while the theoretical capacity of deep architectures is high, the practical expressive power achieved through successful training often falls short. Building on the insights gained from Neural ODEs, which explore the depth of DNNs as a continuous variable, in this work, we generalize the traditional fully connected DNN through the concept of continuous width. In the Generalized Deep Neural Network (GDNN), the traditional notion of neurons in each layer is replaced by a continuous state function. Using the finite rank parameterization of the weight integral kernel, we establish that GDNN can be obtained by employing the Integral Activation Transform (IAT) as activation layers within the traditional DNN framework. The IAT maps the input vector to a function space using some basis functions, followed by nonlinear activation in the function space, and then extracts information through the integration with another collection of basis functions. A specific variant, IAT-ReLU, featuring the ReLU nonlinearity, serves as a smooth generalization of the scalar ReLU activation. Notably, IAT-ReLU exhibits a continuous activation pattern when continuous basis functions are employed, making it smooth and enhancing the trainability of the DNN. Our numerical experiments demonstrate that IAT-ReLU outperforms regular ReLU in terms of trainability and better smoothness.

📄 PDF Abstract BibTeX arXiv:2312.12578

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the expressive power of message-passing neural networks as global feature map transformers

2022-03-17 · Floris Geerts, Jasper Steegmans, Jan Van den Bussche

We investigate the power of message-passing neural networks (MPNNs) in their capacity to transform the numerical features stored in the nodes of their input graphs. Our focus is on global expressive power, uniformly over…

A Logical View of GNN-Style Computation and the Role of Activation Functions

2025-12-22 · Pablo Barceló, Floris Geerts, Matthias Lanzinger, Klara Pakhomenko 외 arxiv

We study the numerical and Boolean expressiveness of MPLang, a declarative language that captures the computation of graph neural networks (GNNs) through linear message passing and activation functions. We begin with A-M…

Activation degree thresholds and expressiveness of polynomial neural networks

2024-08-08 · Bella Finkel, Jose Israel Rodriguez, Chenxi Wu, Thomas Yahl

We study the expressive power of deep polynomial neural networks through the geometry of their neurovariety. We introduce the notion of the activation degree threshold of a network architecture to express when the dimens…

Polynomial Neural Networks

Symmetric-APL Activations: Training Insights and Robustness to Adversarial Attacks

2019-09-25 · Mohammadamin Tavakoli, Forest Agostinelli, Pierre Baldi

Deep neural networks with learnable activation functions have shown superior performance over deep neural networks with fixed activation functions for many different problems. The adaptability of learnable activation fun…

Integral representations of shallow neural network with Rectified Power Unit activation function

2021-12-20 · Ahmed Abdeljawad, Philipp Grohs

In this effort, we derive a formula for the integral representation of a shallow neural network with the Rectified Power Unit activation function. Mainly, our first result deals with the univariate case of representation…