paper-with-me

홈 › Papers

Cheap Orthogonal Constraints in Neural Networks: A Simple Parametrization of the Orthogonal and Unitary Group

2019-01-24 · Mario Lezcano-Casado, David Martínez-Rubio

We introduce a novel approach to perform first-order optimization with orthogonal and unitary constraints. This approach is based on a parametrization stemming from Lie group theory through the exponential map. The parametrization transforms the constrained optimization problem into an unconstrained one over a Euclidean space, for which common first-order optimization methods can be used. The theoretical results presented are general enough to cover the special orthogonal group, the unitary group and, in general, any connected compact Lie group. We discuss how this and other parametrizations can be computed efficiently through an implementation trick, making numerically complex parametrizations usable at a negligible runtime cost in neural networks. In particular, we apply our results to RNNs with orthogonal recurrent weights, yielding a new architecture called expRNN. We demonstrate how our method constitutes a more robust approach to optimization with orthogonal constraints, showing faster, accurate, and more stable convergence in several tasks designed to test RNNs.

📄 PDF Abstract BibTeX arXiv:1901.08428

Code (3)

Lezcano/expRNN 공식 구현 pytorch
Lezcano/geotorch pytorch
ndminhkhoi46/asRNN pytorch

Similar Papers 제목 키워드 기반

Group and Shuffle: Efficient Structured Orthogonal Parametrization

2024-06-14 · Mikhail Gorbunov, Nikolay Yudin, Vera Soboleva, Aibek Alanov 외

The increasing size of neural networks has led to a growing demand for methods of efficient fine-tuning. Recently, an orthogonal fine-tuning paradigm was introduced that uses orthogonal matrices for adapting the weights …

Computational EfficiencyLanguage ModelingLanguage Modelling

Infeasible Deterministic, Stochastic, and Variance-Reduction Algorithms for Optimization under Orthogonality Constraints

2023-03-29 · Pierre Ablin, Simon Vary, Bin Gao, P. -A. Absil

Orthogonality constraints naturally appear in many machine learning problems, from principal component analysis to robust neural network training. They are usually solved using Riemannian optimization algorithms, which m…

Riemannian optimization

O-ViT: Orthogonal Vision Transformer

2022-01-28 · Yanhong Fei, Yingjie Liu, Xian Wei, Mingsong Chen

Inspired by the tremendous success of the self-attention mechanism in natural language processing, the Vision Transformer (ViT) creatively applies it to image patch sequences and achieves incredible performance. However,…

Parallelized Computation and Backpropagation Under Angle-Parametrized Orthogonal Matrices

2021-05-30 · Firas Hamze

We present a methodology for parallel acceleration of learning in the presence of matrix orthogonality and unitarity constraints of interest in several branches of machine learning. We show how an apparently sequential e…

GPU

An Embarrassingly Simple Way to Optimize Orthogonal Matrices at Scale

2026-02-16 · Adrián Javaloy, Antonio Vergari arxiv

Orthogonality constraints are ubiquitous in robust and probabilistic machine learning. Unfortunately, current optimizers are computationally expensive and do not scale to problems with hundreds or thousands of constraint…