paper-with-me

홈 › Papers

Training Invertible Linear Layers through Rank-One Perturbations

2020-10-14 · Andreas Krämer, Jonas Köhler, Frank Noé

Many types of neural network layers rely on matrix properties such as invertibility or orthogonality. Retaining such properties during optimization with gradient-based stochastic optimizers is a challenging task, which is usually addressed by either reparameterization of the affected parameters or by directly optimizing on the manifold. This work presents a novel approach for training invertible linear layers. In lieu of directly optimizing the network parameters, we train rank-one perturbations and add them to the actual weight matrices infrequently. This P$^{4}$Inv update allows keeping track of inverses and determinants without ever explicitly computing them. We show how such invertible blocks improve the mixing and thus the mode separation of the resulting normalizing flows. Furthermore, we outline how the P$^4$ concept can be utilized to retain properties other than invertibility.

📄 PDF Abstract BibTeX arXiv:2010.07033

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ButterflyFlow: Building Invertible Layers with Butterfly Matrices

2022-09-28 · Chenlin Meng, Linqi Zhou, Kristy Choi, Tri Dao 외

Normalizing flows model complex probability distributions using maps obtained by composing invertible layers. Special linear layers such as masked and 1x1 convolutions play a key role in existing architectures because th…

Density Estimation

Symmetric block-low-rank layers for fully reversible multilevel neural networks

2019-12-14 · Bas Peters, Eldad Haber, Keegan Lensink

Factors that limit the size of the input and output of a neural network include memory requirements for the network states/activations to compute gradients, as well as memory for the convolutional kernels or other weight…

Video SegmentationVideo Semantic Segmentation

CDFlow: Building Invertible Layers with Circulant and Diagonal Matrices

2025-10-29 · Xuchen Feng, Siyu Liao arxiv

Normalizing flows are deep generative models that enable efficient likelihood estimation and sampling through invertible transformations. A key challenge is to design linear layers that enhance expressiveness while maint…

Density Estimation

Low-Rank Tensor Completion by Approximating the Tensor Average Rank

2021-01-01 · ICCV 2021 10 · Zhanliang Wang, Junyu Dong, Xinguo Liu, Xueying Zeng

This paper focuses on the problem of low-rank tensor completion, the goal of which is to recover an underlying low-rank tensor from incomplete observations. Our method is motivated by the recently proposed t-product …

Fourier-Invertible Neural Encoder (FINE) for Homogeneous Flows

2025-05-21 · Anqiao Ouyang, Hongyi Ke, Qi Wang

Invertible neural architectures have recently attracted attention for their compactness, interpretability, and information-preserving properties. In this work, we propose the Fourier-Invertible Neural Encoder (FINE), whi…

Dimensionality ReductionPhysics-informed machine learningRepresentation Learning