paper-with-me

Papers

Neuron Campaign for Initialization Guided by Information Bottleneck Theory

2021-08-14 · Haitao Mao, Xu Chen, Qiang Fu, Lun Du, Shi Han, Dongmei Zhang

Initialization plays a critical role in the training of deep neural networks (DNN). Existing initialization strategies mainly focus on stabilizing the training process to mitigate gradient vanish/explosion problems. However, these initialization methods are lacking in consideration about how to enhance generalization ability. The Information Bottleneck (IB) theory is a well-known understanding framework to provide an explanation about the generalization of DNN. Guided by the insights provided by IB theory, we design two criteria for better initializing DNN. And we further design a neuron campaign initialization algorithm to efficiently select a good initialization for a neural network on a given dataset. The experiments on MNIST dataset show that our method can lead to a better generalization performance with faster convergence.

📄 PDF Abstract BibTeX arXiv:2108.06530

Code (1)

huanhuqueyue/cikm-ibci 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Nonlinear scaling of resource allocation in sensory bottlenecks

2019-12-01 · NeurIPS 2019 12 · Laura Rose Edmondson, Alejandro Jimenez Rodriguez, Hannes P. Saal

In many sensory systems, information transmission is constrained by a bottleneck, where the number of output neurons is vastly smaller than the number of input neurons. Efficient coding theory predicts that in these scen…

Persistent Neurons

2020-07-02 · Yimeng Min

Neural networks (NN)-based learning algorithms are strongly affected by the choices of initialization and data distribution. Different optimization strategies have been proposed for improving the learning trajectory and …

Accelerating Training of Deep Spiking Neural Networks with Parameter Initialization

2021-09-29 · Jianhao Ding, Jiyuan Zhang, Zhaofei Yu, Tiejun Huang

Despite that spiking neural networks (SNNs) show strong advantages in information encoding, power consuming, and computational capability, the underdevelopment of supervised learning algorithms is still a hindrance for t…

Attention-Based Guided Structured Sparsity of Deep Neural Networks

2018-02-13 · Amirsina Torfi, Rouzbeh A. Shirvani, Sobhan Soleymani, Nasser M. Nasrabadi

Network pruning is aimed at imposing sparsity in a neural network architecture by increasing the portion of zero-valued weights for reducing its size regarding energy-efficiency consideration and increasing evaluation sp…

Network Pruning

Novel Architectures for Unsupervised Information Bottleneck based Speaker Diarization of Meetings

2020-10-13

Speaker diarization is an important problem that is topical, and is especially useful as a preprocessor for conversational speech related applications. The objective of this paper is two-fold: (i) segment initialization …

Clusteringspeaker-diarizationSpeaker Diarization