paper-with-me

홈 › Papers

Non Vanishing Gradients for Arbitrarily Deep Neural Networks: a Hamiltonian System Approach

2021-09-27 · NeurIPS Workshop DLDE 2021 12 · Clara Galimberti, Luca Furieri, Liang Xu, Giancarlo Ferrari-Trecate

Deep Neural Networks (DNNs) training can be difficult due to vanishing or exploding gradients during weight optimization through backpropagation. To address this problem, we propose a general class of Hamiltonian DNNs (H-DNNs) that stems from the discretization of continuous-time Hamiltonian systems. Our main result is that a broad set of H-DNNs ensures non-vanishing gradients by design for an arbitrary network depth. This is obtained by proving that, using a semi-implicit Euler discretization scheme, the backward sensitivity matrices involved in gradient computations are symplectic.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Sensitivity

Similar Papers 제목 키워드 기반

Hamiltonian Deep Neural Networks Guaranteeing Non-vanishing Gradients by Design

2021-05-27 · Clara Lucía Galimberti, Luca Furieri, Liang Xu, Giancarlo Ferrari-Trecate

Deep Neural Networks (DNNs) training can be difficult due to vanishing and exploding gradients during weight optimization through backpropagation. To address this problem, we propose a general class of Hamiltonian DNNs (…

image-classificationImage ClassificationSensitivity

Distributed neural network control with dependability guarantees: a compositional port-Hamiltonian approach

2021-12-16 · Luca Furieri, Clara Lucía Galimberti, Muhammad Zakwan, Giancarlo Ferrari-Trecate

Large-scale cyber-physical systems require that control policies are distributed, that is, that they only rely on local real-time measurements and communication with neighboring agents. Optimal Distributed Control (ODC) …

A unified framework for Hamiltonian deep neural networks

2021-04-27 · Clara L. Galimberti, Liang Xu, Giancarlo Ferrari Trecate

Training deep neural networks (DNNs) can be difficult due to the occurrence of vanishing/exploding gradients during weight optimization. To avoid this problem, we propose a class of DNNs stemming from the time discretiza…

Improving Parameter Training for VQEs by Sequential Hamiltonian Assembly

2023-12-09 · Jonas Stein, Navid Roshani, Maximilian Zorn, Philipp Altmann 외

A central challenge in quantum machine learning is the design and training of parameterized quantum circuits (PQCs). Similar to deep learning, vanishing gradients pose immense problems in the trainability of PQCs, which …

Quantum Machine Learning

Universal Approximation Property of Hamiltonian Deep Neural Networks

2023-03-21 · Muhammad Zakwan, Massimiliano d'Angelo, Giancarlo Ferrari-Trecate

This paper investigates the universal approximation capabilities of Hamiltonian Deep Neural Networks (HDNNs) that arise from the discretization of Hamiltonian Neural Ordinary Differential Equations. Recently, it has been…