paper-with-me

Papers

High-dimensional Asymptotics of Generalization Performance in Continual Ridge Regression

2025-08-21 · Yihan Zhao, Wenqing Su, Ying Yang arxiv

Continual learning is motivated by the need to adapt to real-world dynamics in tasks and data distribution while mitigating catastrophic forgetting. Despite significant advances in continual learning techniques, the theoretical understanding of their generalization performance lags behind. This paper examines the theoretical properties of continual ridge regression in high-dimensional linear models, where the dimension is proportional to the sample size in each task. Using random matrix theory, we derive exact expressions of the asymptotic prediction risk, thereby enabling the characterization of three evaluation metrics of generalization performance in continual learning: average risk, backward transfer, and forward transfer. Furthermore, we present the theoretical risk curves to illustrate the trends in these evaluation metrics throughout the continual learning process. Our analysis reveals several intriguing phenomena in the risk curves, demonstrating how model specifications influence the generalization performance. Simulation studies are conducted to validate our theoretical findings.

📄 PDF Abstract BibTeX arXiv:2508.15494

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

The Nuclear Route: Sharp Asymptotics of ERM in Overparameterized Quadratic Networks

2025-05-23 · Vittorio Erba, Emanuele Troiani, Lenka Zdeborová, Florent Krzakala

We study the high-dimensional asymptotics of empirical risk minimization (ERM) in over-parametrized two-layer neural networks with quadratic activations trained on synthetic data. We derive sharp asymptotics for both tra…

Asymptotics of Random Feature Regression Beyond the Linear Scaling Regime

2024-03-13 · Hong Hu, Yue M. Lu, Theodor Misiakiewicz

Recent advances in machine learning have been achieved by using overparametrized models trained until near interpolation of the training data. It was shown, e.g., through the double descent phenomenon, that the number of…

regression

The generalization error of max-margin linear classifiers: Benign overfitting and high dimensional asymptotics in the overparametrized regime

2019-11-05 · Andrea Montanari, Feng Ruan, Youngtak Sohn, Jun Yan

Modern machine learning classifiers often exhibit vanishing classification error on the training set. They achieve this by learning nonlinear representations of the inputs that maps the data into linearly separable class…

Optimal L2 Regularization in High-dimensional Continual Linear Regression

2026-01-20 · Gilad Karpel, Edward Moroshko, Ran Levinstein, Ron Meir 외 arxiv

We study generalization in an overparameterized continual linear regression setting, where a model is trained with L2 (isotropic) regularization across a sequence of tasks. We derive a closed-form expression for the expe…

Continual Learning

Asymptotics of Ridge Regression in Convolutional Models

2021-03-08 · Mojtaba Sahraee-Ardakan, Tung Mai, Anup Rao, Ryan Rossi 외

Understanding generalization and estimation error of estimators for simple models such as linear and generalized linear models has attracted a lot of attention recently. This is in part due to an interesting observation …

regression