paper-with-me

홈 › Papers

Retention analysis of edited knowledge after fine-tuning

2025-07-14 · Fufang Wen, Shichang Zhang arxiv

Large language models (LLMs) store vast amounts of knowledge, which often requires updates to correct factual errors, incorporate newly acquired information, or adapt model behavior. Model editing methods have emerged as efficient solutions for such updates, offering localized and precise knowledge modification at significantly lower computational cost than continual training. In parallel, LLMs are frequently fine-tuned for a wide range of downstream tasks. However, the effect of fine-tuning on previously edited knowledge remains poorly understood. In this work, we systematically investigate how different fine-tuning objectives interact with various model editing techniques. Our findings show that edited knowledge is substantially more susceptible to forgetting during fine-tuning than intrinsic knowledge acquired through pre-training. This analysis highlights a key limitation of current editing approaches and suggests that evaluating edit robustness under downstream fine-tuning is critical for their practical deployment. We further find that knowledge retention can be significantly improved by either augmenting edit knowledge with paraphrases or by freezing layers associated with edited content in fine-tuning stage, offering insight for developing more robust editing algorithms.

📄 PDF Abstract BibTeX arXiv:2507.14198

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation

2025-11-08 · Yinjie Cheng, Paul Youssef, Christin Seifert, Jörg Schlötterer 외 arxiv

Knowledge editing (KE) offers a lightweight alternative to retraining for updating large language models (LLMs). Meanwhile, fine-tuning remains the default operation for adapting LLMs to new domains and tasks. Despite th…

knowledge editing

MAML-CL: Edited Model-Agnostic Meta-Learning for Continual Learning

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Continual learning (CL) exhibits a learning ability to well-learn all sequentially seen tasks drawn from various domains. Yet, existing sequential training methods fail to consolidate learned knowledge from earlier tasks…

Continual LearningMeta-Learningtext-classificationText Classification

Uncovering Entity Identity Confusion in Multimodal Knowledge Editing

2026-05-07 · Shu Wu, Xiaotian Ye, Xinyu Mou, Dongsheng Liu 외 arxiv

Multimodal knowledge editing (MKE) aims to correct the internal knowledge of large vision-language models after deployment, yet the behavioral patterns of post-edit models remain underexplored. In this paper, we identify…

knowledge editing

Evidence-State Rewards for Long-Context Reasoning

2026-07-02 · Ya Gao, Pekka Marttinen arxiv

Long-context reasoning requires models to locate, revise, and synthesize evidence distributed across lengthy inputs. Existing long-context RL methods usually reward final answers or static evidence extraction, offering l…

Reinforcement Learning

Beyond Perplexity: A Lightweight Benchmark for Knowledge Retention in Supervised Fine-Tuning

2026-01-07 · Soheil Zibakhsh Shabgahi, Pedram Aghazadeh, Farinaz Koushanfar arxiv

Supervised Fine-Tuning (SFT) is a standard approach for injecting domain knowledge into Large Language Models (LLMs). However, relying on validation perplexity to monitor training is often insufficient, as it confounds s…