paper-with-me

Papers

MLLM-CL: Continual Learning for Multimodal Large Language Models

2025-06-05 · Hongbo Zhao, Fei Zhu, Rundong Wang, Gaofeng Meng, Zhaoxiang Zhang

Recent Multimodal Large Language Models (MLLMs) excel in vision-language understanding but face challenges in adapting to dynamic real-world scenarios that require continuous integration of new knowledge and skills. While continual learning (CL) offers a potential solution, existing benchmarks and methods suffer from critical limitations. In this paper, we introduce MLLM-CL, a novel benchmark encompassing domain and ability continual learning, where the former focuses on independently and identically distributed (IID) evaluation across evolving mainstream domains, whereas the latter evaluates on non-IID scenarios with emerging model ability. Methodologically, we propose preventing catastrophic interference through parameter isolation, along with an MLLM-based routing mechanism. Extensive experiments demonstrate that our approach can integrate domain-specific knowledge and functional abilities with minimal forgetting, significantly outperforming existing methods.

📄 PDF Abstract BibTeX arXiv:2506.05453

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

Improving Multimodal Large Language Models Using Continual Learning

2024-10-25 · Shikhar Srivastava, Md Yousuf Harun, Robik Shrestha, Christopher Kanan

Generative large language models (LLMs) exhibit impressive capabilities, which can be further augmented by integrating a pre-trained vision model into the original LLM to create a multimodal LLM (MLLM). However, this int…

Continual LearningNatural Language Understanding

Is Our Benchmark Enough? An Analysis of Continual Learning for MLLMs

2026-06-18 · Van-Tuan Tran, Shruthi Gowda, Merim Dzaferagic, Marco Ruffini arxiv

Continual adaptation is essential for multimodal large language models (MLLMs) deployed across evolving domains, but the state-of-the-art MR-LoRA method highly relies on the assumption that a MLLM-based router is necessa…

Continual Learning

Continual-NExT: A Unified Comprehension And Generation Continual Learning Framework

2026-02-20 · Jingyang Qiao, Zhizhong Zhang, Xin Tan, Jingyu Gong 외 arxiv

Dual-to-Dual MLLMs refer to Multimodal Large Language Models, which can enable unified multimodal comprehension and generation through text and image modalities. Although exhibiting strong instantaneous learning and gene…

Continual Learning

CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering

2025-03-01 · CVPR 2025 1 · Tianyu Huai, Jie zhou, Xingjiao Wu, Qin Chen 외

Multimodal large language models (MLLMs) have garnered widespread attention from researchers due to their remarkable understanding and generation capabilities in visual language tasks (e.g., visual question answering). H…

Continual LearningLanguage ModelingLanguage ModellingLarge Language Model+5

Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives

2025-11-23 · Kai Jiang, Siqi Huang, Xiangyu Chen, Jiawei Shao 외 arxiv

Multimodal large language models (MLLMs) deployed on devices must adapt to continuously changing visual scenarios such as variations in background and perspective, to effectively perform complex visual tasks. To investig…

Continual Learning