paper-with-me

홈 › Papers

LoRA-Loop: Closing the Synthetic Replay Cycle for Continual VLM Learning

2025-07-17 · Kaihong Wang, Donghyun Kim, Margrit Betke arxiv

Continual learning for vision-language models has achieved remarkable performance through synthetic replay, where samples are generated using Stable Diffusion to regularize during finetuning and retain knowledge. However, real-world downstream applications often exhibit domain-specific nuances and fine-grained semantics not captured by generators, causing synthetic-replay methods to produce misaligned samples that misguide finetuning and undermine retention of prior knowledge. In this work, we propose a LoRA-enhanced synthetic-replay framework that injects task-specific low-rank adapters into a frozen Stable Diffusion model, efficiently capturing each new task's unique visual and semantic patterns. Specifically, we introduce a two-stage, confidence-based sample selection: we first rank real task data by post-finetuning VLM confidence to focus LoRA finetuning on the most representative examples, then generate synthetic samples and again select them by confidence for distillation. Our approach integrates seamlessly with existing replay pipelines-simply swap in the adapted generator to boost replay fidelity. Extensive experiments on the Multi-domain Task Incremental Learning (MTIL) benchmark show that our method outperforms previous synthetic-replay techniques, achieving an optimal balance among plasticity, stability, and zero-shot capability. These results demonstrate the effectiveness of generator adaptation via LoRA for robust continual learning in VLMs.

📄 PDF Abstract BibTeX arXiv:2507.13568

Code (0)

등록된 구현이 없습니다.

Tasks

Incremental LearningContinual Learning

Similar Papers 제목 키워드 기반

Blockchain-Linked Auditable Decision Management for Telecom/IoT Fraud-Control Requests

2026-07-10 · Saviz Changizi, Nasibeh Mohammadzadeh, Mohammad Shojafar, Rahim Tafazolli arxiv

Telecom fraud-control studies often stop at detector-level classification, but deployment use requires request-level policy resolution, lifecycle traceability, and auditability. This paper reframes fraud control as block…

Closing the Loop: Joint Rain Generation and Removal via Disentangled Image Translation

2021-03-25 · CVPR 2021 1 · Yuntong Ye, Yi Chang, Hanyu Zhou, Luxin Yan

Existing deep learning-based image deraining methods have achieved promising performance for synthetic rainy images, typically rely on the pairs of sharp images and simulated rainy counterparts. However, these methods su…

DisentanglementRain RemovalTranslation

Embodied Science: Closing the Discovery Loop with Agentic Embodied AI

2026-03-20 · Xiang Zhuang, Chenyi Zhou, Kehua Feng, Zhihui Zhu 외 arxiv

Artificial intelligence has demonstrated remarkable capability in predicting scientific properties, yet scientific discovery remains an inherently physical, long-horizon pursuit governed by experimental cycles. Most curr…

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

2026-07-02 · Fangfei Li, Chenyang Zhao, Long Wang, Feng Tian 외 arxiv

Domain agents often face noisy business data, uncertain post-training gains, offline/application mismatch, and adapter-release risk. This paper presents CLAP (Closed-Loop Agent Post-training), a closed-loop method that c…

Real-time experiment-theory closed-loop interaction for autonomous materials science

2024-10-22 · Haotong Liang, Chuangye Wang, Heshan Yu, Dylan Kirsch 외

Iterative cycles of theoretical prediction and experimental validation are the cornerstone of the modern scientific method. However, the proverbial "closing of the loop" in experiment-theory cycles in practice are usuall…