paper-with-me

홈 › Papers

DAM: Dynamic Adapter Merging for Continual Video QA Learning

2024-03-13 · Feng Cheng, Ziyang Wang, Yi-Lin Sung, Yan-Bo Lin, Mohit Bansal, Gedas Bertasius

We present a parameter-efficient method for continual video question-answering (VidQA) learning. Our method, named DAM, uses the proposed Dynamic Adapter Merging to (i) mitigate catastrophic forgetting, (ii) enable efficient adaptation to continually arriving datasets, (iii) handle inputs from unknown datasets during inference, and (iv) enable knowledge sharing across similar dataset domains. Given a set of continually streaming VidQA datasets, we sequentially train dataset-specific adapters for each dataset while freezing the parameters of a large pretrained video-language backbone. During inference, given a video-question sample from an unknown domain, our method first uses the proposed non-parametric router function to compute a probability for each adapter, reflecting how relevant that adapter is to the current video-question input instance. Subsequently, the proposed dynamic adapter merging scheme aggregates all the adapter weights into a new adapter instance tailored for that particular test sample to compute the final VidQA prediction, mitigating the impact of inaccurate router predictions and facilitating knowledge sharing across domains. Our DAM model outperforms prior state-of-the-art continual learning approaches by 9.1% while exhibiting 1.9% less forgetting on 6 VidQA datasets spanning various domains. We further extend DAM to continual image classification and image QA and outperform prior methods by a large margin. The code is publicly available at: https://github.com/klauscc/DAM

📄 PDF Abstract BibTeX arXiv:2403.08755

Code (1)

klauscc/dam 공식 구현 pytorch

Tasks

Continual Learningimage-classificationImage ClassificationQuestion AnsweringVideo Question Answering

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Adapter 설명 없음

Similar Papers 제목 키워드 기반

Beyond Classification: Dynamic Adapter Routing for Continual Multimodal Retrieval

2026-05-29 · Alicja Dobrzeniecka, Filip Szatkowski, Sebastian Cygert, Szymon Lukasik 외 arxiv

While retrieval is a core function of vision-language models, continually updating these models for retrieval tasks remains critically underexplored. Existing work often approaches continual retrieval through the lens of…

class-incremental learning

HAM: Hierarchical Adapter Merging for Scalable Continual Learning

2025-09-16 · Eric Nuertey Coleman, Luigi Quarantiello, Samrat Mukherjee, Julio Hurtado 외 arxiv

Continual learning is an essential capability of human cognition, yet it poses significant challenges for current deep learning models. The primary issue is that new knowledge can interfere with previously learned inform…

parameter-efficient fine-tuningContinual LearningTransfer Learning

Dual-Imbalance Continual Learning for Real-World Food Recognition

2026-03-31 · Xiaoyan Zhang, Jiangpeng He arxiv

Visual food recognition in real-world dietary logging scenarios naturally exhibits severe data imbalance, where a small number of food categories appear frequently while many others occur rarely, resulting in long-tailed…

parameter-efficient fine-tuningIncremental LearningContinual Learning

Low-Rank Continual Personalization of Diffusion Models

2024-10-07 · Łukasz Staniszewski, Katarzyna Zaleska, Kamil Deja

Recent personalization methods for diffusion models, such as Dreambooth, allow fine-tuning pre-trained models to generate new concepts. However, applying these techniques across multiple tasks in order to include, e.g., …

Continual Learning

RegCL: Continual Adaptation of Segment Anything Model via Model Merging

2025-07-16 · Yuan-Chen Shu, Zhiwei Lin, Yongtao Wang

To address the performance limitations of the Segment Anything Model (SAM) in specific domains, existing works primarily adopt adapter-based one-step adaptation paradigms. However, some of these methods are specific deve…

Continual Learningmodel