paper-with-me

홈 › Papers

Dynamic Transformer Architecture for Continual Learning of Multimodal Tasks

2024-01-27 · Yuliang Cai, Mohammad Rostami

Transformer neural networks are increasingly replacing prior architectures in a wide range of applications in different data modalities. The increasing size and computational demands of fine-tuning large pre-trained transformer neural networks pose significant challenges for the widespread adoption of these models for applications that demand on-edge computing. To tackle this challenge, continual learning (CL) emerges as a solution by facilitating the transfer of knowledge across tasks that arrive sequentially for an autonomously learning agent. However, current CL methods mainly focus on learning tasks that are exclusively vision-based or language-based. We propose a transformer-based CL framework focusing on learning tasks that involve both vision and language, known as Vision-and-Language (VaL) tasks. Due to the success of transformers in other modalities, our architecture has the potential to be used in multimodal learning settings. In our framework, we benefit from introducing extra parameters to a base transformer to specialize the network for each task. As a result, we enable dynamic model expansion to learn several tasks in a sequence. We also use knowledge distillation to benefit from relevant past experiences to learn the current task more efficiently. Our proposed method, Task Attentive Multimodal Continual Learning (TAM-CL), allows for the exchange of information between tasks while mitigating the problem of catastrophic forgetting. Notably, our approach is scalable, incurring minimal memory and time overhead. TAM-CL achieves state-of-the-art (SOTA) performance on challenging multimodal tasks

📄 PDF Abstract BibTeX arXiv:2401.15275

Code (0)

등록된 구현이 없습니다.

Tasks

Continual LearningEdge-computingKnowledge Distillation

Methods 이 논문이 사용한 방법론

Focus 설명 없음
Knowledge Distillation A very simple way to improve the performance of almost any machine learning algorithm is to train many different models on the same data and then to average their predictions.…
BASE 설명 없음

Similar Papers 제목 키워드 기반

Dynamic Mixture of Curriculum LoRA Experts for Continual Multimodal Instruction Tuning

2025-06-13 · Chendi Ge, Xin Wang, Zeyang Zhang, Hong Chen 외

Continual multimodal instruction tuning is crucial for adapting Multimodal Large Language Models (MLLMs) to evolving tasks. However, most existing methods adopt a fixed architecture, struggling with adapting to new tasks…

Continual Learning

Multimodal Continual Instruction Tuning with Dynamic Gradient Guidance

2025-11-19 · Songze Li, Mingyu Gao, Tonghua Su, Xu-Yao Zhang 외 arxiv

Multimodal continual instruction tuning enables multimodal large language models to sequentially adapt to new tasks while building upon previously acquired knowledge. However, this continual learning paradigm faces the s…

Continual Learning

DyTox: Transformers for Continual Learning with DYnamic TOken eXpansion

2021-11-22 · CVPR 2022 1 · Arthur Douillard, Alexandre Ramé, Guillaume Couairon, Matthieu Cord

Deep network architectures struggle to continually learn new tasks without forgetting the previous tasks. A recent trend indicates that dynamic architectures based on an expansion of the parameters can reduce catastrophi…

Class Incremental LearningContinual LearningDecoderIncremental Learning

CLiMB: A Continual Learning Benchmark for Vision-and-Language Tasks

2022-06-18 · Tejas Srinivasan, Ting-Yun Chang, Leticia Leonor Pinto Alva, Georgios Chochlakis 외

Current state-of-the-art vision-and-language models are evaluated on tasks either individually or in a multi-task setting, overlooking the challenges of continually learning (CL) tasks as they arrive. Existing CL benchma…

Continual LearningTransfer Learning

Dynamic Dialogue Policy for Continual Reinforcement Learning

2022-04-12 · COLING 2022 10 · Christian Geishauser, Carel van Niekerk, Nurul Lubis, Michael Heck 외

Continual learning is one of the key components of human learning and a necessary requirement of artificial intelligence. As dialogue can potentially span infinitely many topics and tasks, a task-oriented dialogue system…

Continual Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)