paper-with-me

홈 › Papers

Machine Unlearning on Pre-trained Models by Residual Feature Alignment Using LoRA

2024-11-13 · Laiqiao Qin, Tianqing Zhu, LinLin Wang, Wanlei Zhou

Machine unlearning is new emerged technology that removes a subset of the training data from a trained model without affecting the model performance on the remaining data. This topic is becoming increasingly important in protecting user privacy and eliminating harmful or outdated data. The key challenge lies in effectively and efficiently unlearning specific information without compromising the model's utility on the retained data. For the pre-trained models, fine-tuning is an important way to achieve the unlearning target. Previous work typically fine-tuned the entire model's parameters, which incurs significant computation costs. In addition, the fine-tuning process may cause shifts in the intermediate layer features, affecting the model's overall utility. In this work, we propose a novel and efficient machine unlearning method on pre-trained models. We term the method as Residual Feature Alignment Unlearning. Specifically, we leverage LoRA (Low-Rank Adaptation) to decompose the model's intermediate features into pre-trained features and residual features. By adjusting the residual features, we align the unlearned model with the pre-trained model at the intermediate feature level to achieve both unlearning and remaining targets. The method aims to learn the zero residuals on the retained set and shifted residuals on the unlearning set. Extensive experiments on numerous datasets validate the effectiveness of our approach.

📄 PDF Abstract BibTeX arXiv:2411.08443

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Unlearning

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

MeGU: Machine-Guided Unlearning with Target Feature Disentanglement

2026-02-19 · Haoyu Wang, Zhuo Huang, Xiaolong Wang, Bo Han 외 arxiv

The growing concern over training data privacy has elevated the "Right to be Forgotten" into a critical requirement, thereby raising the demand for effective Machine Unlearning. However, existing unlearning approaches co…

Revisiting Machine Unlearning with Dimensional Alignment

2024-07-25 · Seonguk Seo, Dongwan Kim, Bohyung Han

Machine unlearning, an emerging research topic focusing on compliance with data privacy regulations, enables trained models to remove the information learned from specific data. While many existing methods indirectly add…

Machine Unlearning

The Unseen Threat: Residual Knowledge in Machine Unlearning under Perturbed Samples

2026-01-29 · Hsiang Hsu, Pradeep Niroula, Zichang He, Ivan Brugere 외 arxiv

Machine unlearning offers a practical alternative to avoid full model re-training by approximately removing the influence of specific user data. While existing methods certify unlearning via statistical indistinguishabil…

Machine Unlearning Method Based On Projection Residual

2022-09-30 · Zihao Cao, Jianzong Wang, Shijing Si, Zhangcheng Huang 외

Machine learning models (mainly neural networks) are used more and more in real life. Users feed their data to the model for training. But these processes are often one-way. Once trained, the model remembers the data. Ev…

Machine Unlearning

Reminiscence Attack on Residuals: Exploiting Approximate Machine Unlearning for Privacy

2025-07-28 · Yaxin Xiao, Qingqing Ye, Li Hu, Huadi Zheng 외 arxiv

Machine unlearning enables the removal of specific data from ML models to uphold the right to be forgotten. While approximate unlearning algorithms offer efficient alternatives to full retraining, this work reveals that …