paper-with-me

홈 › Papers

The role of class encoding in neural collapse

2026-05-29 · Bastien Massion, Roy Makhlouf, Estelle Massart arxiv

Neural collapse is a structural property of the last-hidden-layer activations in neural network classification models, when trained beyond a zero classification error. In this work, we explore the role of label encoding in neural collapse by relying on the unrestricted feature model with mean squared error training loss. We demonstrate that, for one-hot encoded labels and balanced data, the uncentered mean features associated with each class transition from a simplex equiangular tight frame to an orthogonal frame when increasing the bias regularization coefficient associated with the final classifier. These structures are reminiscent of the orthogonal frame structure of one-hot encoded labels. For any arbitrary encoding, we also show that the final classifier's bias aims at centering the labels, compensating for the discrepancy between the global mean of the labels and the origin. We further discuss the role of the encoding in other neural collapse properties.

📄 PDF Abstract BibTeX arXiv:2606.00344

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Anti Mode-Collapse in Mean-Field Transformer via Auxiliary Variables

2026-05-28 · Masaaki Imaizumi, Masanori Koyama, Noboru Isobe, Kohei Hayashi arxiv

We use a mean-field-based transformer model to theoretically investigate how auxiliary variables, such as positional encoding, prevent mode collapse of self-attention mechanisms. The use of mean-field transformers to ana…

Limitations of Neural Collapse for Understanding Generalization in Deep Learning

2022-02-17 · Like Hui, Mikhail Belkin, Preetum Nakkiran

The recent work of Papyan, Han, & Donoho (2020) presented an intriguing "Neural Collapse" phenomenon, showing a structural property of interpolating classifiers in the late stage of training. This opened a rich area of e…

Deep LearningRepresentation Learning

Emergence of Latent Binary Encoding in Deep Neural Network Classifiers

2023-10-12 · Luigi Sbailò, Luca Ghiringhelli

We investigate the emergence of binary encoding within the latent space of deep-neural-network classifiers. Such binary encoding is induced by the introduction of a linear penultimate layer, which employs during training…

Feature Collapse

2023-05-25 · Thomas Laurent, James H. von Brecht, Xavier Bresson

We formalize and study a phenomenon called feature collapse that makes precise the intuitive idea that entities playing a similar role in a learning task receive similar representations. As feature collapse requires a no…

On the Role of Attention Masks and LayerNorm in Transformers

2024-05-29 · Xinyi Wu, Amir Ajorlou, Yifei Wang, Stefanie Jegelka 외

Self-attention is the key mechanism of transformers, which are the essential building blocks of modern foundation models. Recent studies have shown that pure self-attention suffers from an increasing degree of rank colla…