paper-with-me

홈 › Papers

Fast Unlearning at Scale via Margin Self-Correction

2026-06-01 · Federico Di Gennaro, Alexander Shevchenko, Fanny Yang arxiv

Language-model unlearning updates a trained model to behave as if it had not seen selected training examples, while preserving utility and avoiding costly retraining. Existing approaches typically fine-tune the pretrained model with a fixed training budget and select the final model afterwards by evaluating several saved checkpoints on downstream validation data. Two sources of unnecessary computation limit scalability: training beyond the desired forget-retain trade-off, and checkpoint selection that requires extra storage and repeated evaluations. To address these limitations, we introduce MArgin Self-Correction (MASC), an efficient unlearning method with an online stopping rule that does not require downstream evaluation. Given a text sequence to be forgotten, MASC actively reduces the logit gap between the original next token and the most likely alternatives. It outputs a final model once this gap is small on average over a sufficiently large proportion of token positions across all forget sequences. On TOFU, MUSE News, and MUSE Books, MASC achieves a competitive forget-retain trade-off at a fraction of the computational cost of existing baselines. We further observe that as we increase model size (a.k.a. number of parameters), the trade-offs improve for both MASC and SimNPO -- the forget metrics remain comparable while retain utility increases.

📄 PDF Abstract BibTeX arXiv:2606.02920

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Approximate Machine Unlearning through Manifold Representation Forgetting Guided by Self Mode Connectivity

2026-05-20 · Weiqi Wang, Zhiyi Tian, Chenhan Zhang, Luoyu Chen 외 arxiv

Machine unlearning is a fundamental mechanism that enforces the right to be forgotten. Existing unlearning studies that rely on label manipulation or task-gradient reversal often deliver limited unlearning effectiveness.…

Semantic Similarity

UNO: Unlearning via Orthogonalization in Generative models

2025-06-05 · Pinak Mandal, Georg A. Gottwald

As generative models become increasingly powerful and pervasive, the ability to unlearn specific data, whether due to privacy concerns, legal requirements, or the correction of harmful content, has become increasingly im…

Offset Unlearning for Large Language Models

2024-04-17 · James Y. Huang, Wenxuan Zhou, Fei Wang, Fred Morstatter 외

Despite the strong capabilities of Large Language Models (LLMs) to acquire knowledge from their training corpora, the memorization of sensitive information in the corpora such as copyrighted, harmful, and private content…

Memorization

GDGU: A Gradient Difference-based Graph Unlearning Method for Cyberattack Localization in Electric Vehicle Charging Networks

2026-06-17 · Nanhong Liu, Mucun Sun, Jie Zhang arxiv

Electric vehicle charging stations (EVCSs) can expose distribution feeders to cyberattacks. While machine learning methods, including graph neural networks, can localize which bus is compromised, significant challenges r…

Multi-Label ClassificationGraph Neural Network

A Comprehensive Evaluation of LLM Unlearning Robustness under Multi-Turn Interaction

2026-02-28 · Ruihao Pan, Suhang Wang arxiv

Machine unlearning aims to remove the influence of specific training data from pre-trained models without retraining from scratch, and is increasingly important for large language models (LLMs) due to safety, privacy, an…