paper-with-me

홈 › Papers

Hamiltonian Deep Neural Networks Guaranteeing Non-vanishing Gradients by Design

2021-05-27 · Clara Lucía Galimberti, Luca Furieri, Liang Xu, Giancarlo Ferrari-Trecate

Deep Neural Networks (DNNs) training can be difficult due to vanishing and exploding gradients during weight optimization through backpropagation. To address this problem, we propose a general class of Hamiltonian DNNs (H-DNNs) that stem from the discretization of continuous-time Hamiltonian systems and include several existing DNN architectures based on ordinary differential equations. Our main result is that a broad set of H-DNNs ensures non-vanishing gradients by design for an arbitrary network depth. This is obtained by proving that, using a semi-implicit Euler discretization scheme, the backward sensitivity matrices involved in gradient computations are symplectic. We also provide an upper-bound to the magnitude of sensitivity matrices and show that exploding gradients can be controlled through regularization. Finally, we enable distributed implementations of backward and forward propagation algorithms in H-DNNs by characterizing appropriate sparsity constraints on the weight matrices. The good performance of H-DNNs is demonstrated on benchmark classification problems, including image classification with the MNIST dataset.

📄 PDF Abstract BibTeX arXiv:2105.13205

Code (3)

DecodEPFL/HamiltonianNet 공식 구현 pytorch
ClaraGalimberti/HamiltonianNet pytorch
decodepfl/deepdiscoph pytorch

Tasks

image-classificationImage ClassificationSensitivity

Similar Papers 제목 키워드 기반

Non Vanishing Gradients for Arbitrarily Deep Neural Networks: a Hamiltonian System Approach

2021-09-27 · NeurIPS Workshop DLDE 2021 12 · Clara Galimberti, Luca Furieri, Liang Xu, Giancarlo Ferrari-Trecate

Deep Neural Networks (DNNs) training can be difficult due to vanishing or exploding gradients during weight optimization through backpropagation. To address this problem, we propose a general class of Hamiltonian DNNs (…

Sensitivity

Improving Parameter Training for VQEs by Sequential Hamiltonian Assembly

2023-12-09 · Jonas Stein, Navid Roshani, Maximilian Zorn, Philipp Altmann 외

A central challenge in quantum machine learning is the design and training of parameterized quantum circuits (PQCs). Similar to deep learning, vanishing gradients pose immense problems in the trainability of PQCs, which …

Quantum Machine Learning

Distributed neural network control with dependability guarantees: a compositional port-Hamiltonian approach

2021-12-16 · Luca Furieri, Clara Lucía Galimberti, Muhammad Zakwan, Giancarlo Ferrari-Trecate

Large-scale cyber-physical systems require that control policies are distributed, that is, that they only rely on local real-time measurements and communication with neighboring agents. Optimal Distributed Control (ODC) …

Universal Approximation Property of Hamiltonian Deep Neural Networks

2023-03-21 · Muhammad Zakwan, Massimiliano d'Angelo, Giancarlo Ferrari-Trecate

This paper investigates the universal approximation capabilities of Hamiltonian Deep Neural Networks (HDNNs) that arise from the discretization of Hamiltonian Neural Ordinary Differential Equations. Recently, it has been…

A unified framework for Hamiltonian deep neural networks

2021-04-27 · Clara L. Galimberti, Liang Xu, Giancarlo Ferrari Trecate

Training deep neural networks (DNNs) can be difficult due to the occurrence of vanishing/exploding gradients during weight optimization. To avoid this problem, we propose a class of DNNs stemming from the time discretiza…