paper-with-me

Papers

Implicit biases in multitask and continual learning from a backward error analysis perspective

2023-11-01 · Benoit Dherin

Using backward error analysis, we compute implicit training biases in multitask and continual learning settings for neural networks trained with stochastic gradient descent. In particular, we derive modified losses that are implicitly minimized during training. They have three terms: the original loss, accounting for convergence, an implicit flatness regularization term proportional to the learning rate, and a last term, the conflict term, which can theoretically be detrimental to both convergence and implicit regularization. In multitask, the conflict term is a well-known quantity, measuring the gradient alignment between the tasks, while in continual learning the conflict term is a new quantity in deep learning optimization, although a basic tool in differential geometry: The Lie bracket between the task gradients.

📄 PDF Abstract BibTeX arXiv:2311.00235

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

Implicit Gradient Regularization

2020-09-23 · ICLR 2021 1 · David G. T. Barrett, Benoit Dherin

Gradient descent can be surprisingly good at optimizing deep neural networks without overfitting and without explicit regularization. We find that the discrete steps of gradient descent implicitly regularize models by pe…

Linear Mode Connectivity in Multitask and Continual Learning

2020-10-09 · ICLR 2021 1 · Seyed Iman Mirzadeh, Mehrdad Farajtabar, Dilan Gorur, Razvan Pascanu 외

Continual (sequential) training and multitask (simultaneous) training are often attempting to solve the same overall objective: to find a solution that performs well on all considered tasks. The main difference is in the…

Continual LearningLinear Mode Connectivity

A Theory for Knowledge Transfer in Continual Learning

2022-08-14 · Diana Benavides-Prado, Patricia Riddle

Continual learning of a stream of tasks is an active area in deep neural networks. The main challenge investigated has been the phenomenon of catastrophic forgetting or interference of newly acquired knowledge with knowl…

Continual LearningTransfer Learning

On the Implicit Adversariality of Catastrophic Forgetting in Deep Continual Learning

2025-10-10 · Ze Peng, Jian Zhang, Jintao Guo, Lei Qi 외 arxiv

Continual learning seeks the human-like ability to accumulate new skills in machine intelligence. Its central challenge is catastrophic forgetting, whose underlying cause has not been fully understood for deep networks. …

Continual LearningAdversarial Attack

The African Woman is Rhythmic and Soulful: An Investigation of Implicit Biases in LLM Open-ended Text Generation

2024-07-01 · Serene Lim, María Pérez-Ortiz

This paper investigates the subtle and often concealed biases present in Large Language Models (LLMs), focusing on implicit biases that may remain despite passing explicit bias tests. Implicit biases are significant beca…

EthicsText Generation