paper-with-me

홈 › Papers

Perturbation Analysis of Neural Collapse

2022-10-29 · Tom Tirer, Haoxiang Huang, Jonathan Niles-Weed

Training deep neural networks for classification often includes minimizing the training loss beyond the zero training error point. In this phase of training, a "neural collapse" behavior has been observed: the variability of features (outputs of the penultimate layer) of within-class samples decreases and the mean features of different classes approach a certain tight frame structure. Recent works analyze this behavior via idealized unconstrained features models where all the minimizers exhibit exact collapse. However, with practical networks and datasets, the features typically do not reach exact collapse, e.g., because deep layers cannot arbitrarily modify intermediate features that are far from being collapsed. In this paper, we propose a richer model that can capture this phenomenon by forcing the features to stay in the vicinity of a predefined features matrix (e.g., intermediate features). We explore the model in the small vicinity case via perturbation analysis and establish results that cannot be obtained by the previously studied models. For example, we prove reduction in the within-class variability of the optimized features compared to the predefined input features (via analyzing gradient flow on the "central-path" with minimal assumptions), analyze the minimizers in the near-collapse regime, and provide insights on the effect of regularization hyperparameters on the closeness to collapse. We support our theory with experiments in practical deep learning settings.

📄 PDF Abstract BibTeX arXiv:2210.16658

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FG-UAP: Feature-Gathering Universal Adversarial Perturbation

2022-09-27 · Zhixing Ye, Xinwen Cheng, Xiaolin Huang

Deep Neural Networks (DNNs) are susceptible to elaborately designed perturbations, whether such perturbations are dependent or independent of images. The latter one, called Universal Adversarial Perturbation (UAP), is ve…

LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training

2026-05-28 · Minju Gwak, Minseo Kwak, Dongseok Lee, Guijin Son 외 arxiv

Reinforcement learning (RL) post-training has shown to improve reasoning in large language models (LLMs). However, there has been little exploration on the problem of data contamination in RL post-training, potentially u…

Reinforcement Learning

Collapse-Aware Triplet Decoupling for Adversarially Robust Image Retrieval

2023-12-12 · Qiwei Tian, Chenhao Lin, Zhengyu Zhao, Qian Li 외

Adversarial training has achieved substantial performance in defending image retrieval against adversarial examples. However, existing studies in deep metric learning (DML) still suffer from two major limitations: weak a…

Adversarial DefenseImage RetrievalMetric LearningRetrieval+1

Overfitting or Underfitting? Understand Robustness Drop in Adversarial Training

2020-10-15 · Zichao Li, Liyuan Liu, chengyu dong, Jingbo Shang

Our goal is to understand why the robustness drops after conducting adversarial training for too long. Although this phenomenon is commonly explained as overfitting, our analysis suggest that its primary cause is perturb…

A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning

2026-05-04 · Arahan Kujur arxiv

We show that a threshold in decision capacity determines whether self-play reinforcement learning agents collapse under asymmetric rule perturbations. Across poker variants, matrix games, a dice game, and multiple learni…

Reinforcement Learning