paper-with-me

홈 › Papers

Guiding Neural Collapse: Optimising Towards the Nearest Simplex Equiangular Tight Frame

2024-11-02 · Evan Markou, Thalaiyasingam Ajanthan, Stephen Gould

Neural Collapse (NC) is a recently observed phenomenon in neural networks that characterises the solution space of the final classifier layer when trained until zero training loss. Specifically, NC suggests that the final classifier layer converges to a Simplex Equiangular Tight Frame (ETF), which maximally separates the weights corresponding to each class. By duality, the penultimate layer feature means also converge to the same simplex ETF. Since this simple symmetric structure is optimal, our idea is to utilise this property to improve convergence speed. Specifically, we introduce the notion of nearest simplex ETF geometry for the penultimate layer features at any given training iteration, by formulating it as a Riemannian optimisation. Then, at each iteration, the classifier weights are implicitly set to the nearest simplex ETF by solving this inner-optimisation, which is encapsulated within a declarative node to allow backpropagation. Our experiments on synthetic and real-world architectures for classification tasks demonstrate that our approach accelerates convergence and enhances training stability.

📄 PDF Abstract BibTeX arXiv:2411.01248

Code (1)

evanmarkou/Guiding-Neural-Collapse 공식 구현 jax

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Neural Collapse: A Review on Modelling Principles and Generalization

2022-06-08 · Vignesh Kothapalli

Deep classifier neural networks enter the terminal phase of training (TPT) when training error reaches zero and tend to exhibit intriguing Neural Collapse (NC) properties. Neural collapse essentially represents a state a…

Transfer Learning

Prevalence of Neural Collapse during the terminal phase of deep learning training

2020-08-18 · Vardan Papyan, X. Y. Han, David L. Donoho

Modern practice for training classification deepnets involves a Terminal Phase of Training (TPT), which begins at the epoch where training error first vanishes; During TPT, the training error stays effectively zero while…

Inductive Bias

Neural Collapse with Cross-Entropy Loss

2020-12-15 · Jianfeng Lu, Stefan Steinerberger

We consider the variational problem of cross-entropy loss with $n$ feature vectors on a unit hypersphere in $\mathbb{R}^d$. We prove that when $d \geq n - 1$, the global minimum is given by the simplex equiangular tight …

On the Role of Neural Collapse in Meta Learning Models for Few-shot Learning

2023-09-30 · Saaketh Medepalli, Naren Doraiswamy

Meta-learning frameworks for few-shot learning aims to learn models that can learn new skills or adapt to new environments rapidly with a few training examples. This has led to the generalizability of the developed model…

Few-Shot LearningMeta-Learning

Leveraging Intermediate Neural Collapse with Simplex ETFs for Efficient Deep Neural Networks

2024-12-01 · Emily Liu

Neural collapse is a phenomenon observed during the terminal phase of neural network training, characterized by the convergence of network activations, class means, and linear classifier weights to a simplex equiangular …