paper-with-me

홈 › Papers

The Gentle Collapse: Distributional Metrics for Continual Learning

2026-06-23 · Ahmed Anwar, Andreas Wagner, Federico Raue, Tobias Nauen, Andreas Dengel arxiv

Accuracy degradation is the standard metric for Catastrophic Forgetting (CF), however, it records only whether forgetting occurred or not. It saturates at the extremes and collapses discretely at task boundaries, hiding the internal structure of what is being forgotten. We introduce six softmax-derived metrics spanning true-label rank (TLR), predictive confidence, and distributional divergence that characterize forgetting continuously, each normalized to [0, 1] with no modification to training. On CIFAR-100, these metrics carry information where accuracy does not: at 0% accuracy, the Confusion Margin spans an IQR of [0.32, 0.50] across classes that accuracy treats identically. We demonstrate that this richer signal is actionable in mitigating catastrophic forgetting. Per-sample metric scores used as loss weights reduce forgetting by 1.3 percentage points over uniform experience replay (ER) on CIFAR-100. Furthermore, the slope of a metric over a small window provides a stable sampling criterion: at a small-window size (e.g. 3 epochs), accuracy-trend degrades to 34.79% (std. = 2.32) while log-TLR achieves 41.07% (std. = 0.57). This gap is structural since reliable small-window trend estimation requires a continuous signal. On TinyImageNet, log-TLR trend sampling reduces forgetting by 7.7 percentage points over the ER baseline.

📄 PDF Abstract BibTeX arXiv:2606.25165

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

Mr Darcy and Mr Toad, gentlemen: distributional names and their kinds

2015-04-01 · WS 2015 4 · Aur{\'e}lie Herbelot

Sample-optimal learning of quantum states using gentle measurements

2025-05-30 · Cristina Butucea, Jan Johannes, Henning Stein

Gentle measurements of quantum states do not entirely collapse the initial state. Instead, they provide a post-measurement state at a prescribed trace distance $\alpha$ from the initial state together with a random varia…

LEMMA

Evaluating Stochastic Collapse and Implicit Bias in Multimodal Large Language Models

2026-06-04 · Huiyuan Zheng, Houtao Zhang, Boyang Wang, Qingyi Si 외 arxiv

Current evaluations for Multimodal Large Language Models (MLLMs) overwhelmingly focus on utility-driven objectives, leaving model behavior under logic-neutral scenarios largely underexplored. Stochasticity is essential i…

How to Synthesize Text Data without Model Collapse?

2024-12-19 · Xuekai Zhu, Daixuan Cheng, Hengli Li, Kaiyan Zhang 외

Model collapse in synthetic data indicates that iterative training on self-generated data leads to a gradual decline in performance. With the proliferation of AI models, synthetic data will fundamentally reshape the web …

Overcoming Mode Collapse with Adaptive Multi Adversarial Training

2021-12-29 · Karttikeya Mangalam, Rohin Garg

Generative Adversarial Networks (GANs) are a class of generative models used for various applications, but they have been known to suffer from the mode collapse problem, in which some modes of the target distribution are…

Continual Learning