Training Invertible Linear Layers through Rank-One Perturbations
Many types of neural network layers rely on matrix properties such as invertibility or orthogonality. Retaining such properties during optimization with gradient-based stochastic optimizers is a challenging task, which is usually addressed by either reparameterization of the affected parameters or by directly optimizing on the manifold. This work presents a novel approach for training invertible linear layers. In lieu of directly optimizing the network parameters, we train rank-one perturbations and add them to the actual weight matrices infrequently. This P$^{4}$Inv update allows keeping track of inverses and determinants without ever explicitly computing them. We show how such invertible blocks improve the mixing and thus the mode separation of the resulting normalizing flows. Furthermore, we outline how the P$^4$ concept can be utilized to retain properties other than invertibility.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
ButterflyFlow: Building Invertible Layers with Butterfly Matrices
Normalizing flows model complex probability distributions using maps obtained by composing invertible layers. Special linear layers such as masked and 1x1 convolutions play a key role in existing architectures because th…
Density EstimationSymmetric block-low-rank layers for fully reversible multilevel neural networks
Factors that limit the size of the input and output of a neural network include memory requirements for the network states/activations to compute gradients, as well as memory for the convolutional kernels or other weight…
Video SegmentationVideo Semantic SegmentationCDFlow: Building Invertible Layers with Circulant and Diagonal Matrices
Normalizing flows are deep generative models that enable efficient likelihood estimation and sampling through invertible transformations. A key challenge is to design linear layers that enhance expressiveness while maint…
Density EstimationLow-Rank Tensor Completion by Approximating the Tensor Average Rank
This paper focuses on the problem of low-rank tensor completion, the goal of which is to recover an underlying low-rank tensor from incomplete observations. Our method is motivated by the recently proposed t-product …
Fourier-Invertible Neural Encoder (FINE) for Homogeneous Flows
Invertible neural architectures have recently attracted attention for their compactness, interpretability, and information-preserving properties. In this work, we propose the Fourier-Invertible Neural Encoder (FINE), whi…
Dimensionality ReductionPhysics-informed machine learningRepresentation Learning