paper-with-me

홈 › Papers

Rotation Control Unlearning: Quantifying and Controlling Continuous Unlearning for LLM with The Cognitive Rotation Space

2025-09-30 · Xiang Zhang, Kun Wei, Xu Yang, Jiahua Li, Su Yan, Cheng Deng arxiv

As Large Language Models (LLMs) become increasingly prevalent, their security vulnerabilities have already drawn attention. Machine unlearning is introduced to seek to mitigate these risks by removing the influence of undesirable data. However, existing methods not only rely on the retained dataset to preserve model utility, but also suffer from cumulative catastrophic utility loss under continuous unlearning requests. To solve this dilemma, we propose a novel method, called Rotation Control Unlearning (RCU), which leverages the rotational salience weight of RCU to quantify and control the unlearning degree in the continuous unlearning process. The skew symmetric loss is designed to construct the existence of the cognitive rotation space, where the changes of rotational angle can simulate the continuous unlearning process. Furthermore, we design an orthogonal rotation axes regularization to enforce mutually perpendicular rotation directions for continuous unlearning requests, effectively minimizing interference and addressing cumulative catastrophic utility loss. Experiments on multiple datasets confirm that our method without retained dataset achieves SOTA performance.

📄 PDF Abstract BibTeX arXiv:2509.25743

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unlearning with Control: Assessing Real-world Utility for Large Language Model Unlearning

2024-06-13 · Qizhou Wang, Bo Han, Puning Yang, Jianing Zhu 외

The compelling goal of eradicating undesirable data behaviors, while preserving usual model functioning, underscores the significance of machine unlearning within the domain of large language models (LLMs). Recent resear…

Language ModelingLanguage ModellingLarge Language ModelMachine Unlearning

FROC: A Unified Framework with Risk-Optimized Control for Machine Unlearning in LLMs

2025-12-15 · Si Qi Goh, Yongsen Zheng, Ziyao Liu, Sami Hormi 외 arxiv

Machine unlearning (MU) seeks to eliminate the influence of specific training examples from deployed models. As large language models (LLMs) become widely used, managing risks arising from insufficient forgetting or util…

Graph-Guided Selective Unlearning for Language Models: Controlling Support Routes Beyond Forget Seeds

2026-08-27 · Waqas Khan, Tabinda Sarwar, Jingyue Cong, Xun Yi 외 arxiv

Enterprises fine-tune language models on proprietary data that may later require removal due to privacy, contractual, or compliance obligations. Selective unlearning removes requested knowledge while preserving model uti…

ALTER: Asymmetric LoRA for Token-Entropy-Guided Unlearning of LLMs

2026-03-02 · Xunlei Chen, Jinyu Guo, Yuang Li, Zhaokun Wang 외 arxiv

Large language models (LLMs) have advanced to encompass extensive knowledge across diverse domains. Yet controlling what a LLMs should not know is important for ensuring alignment and thus safe use. However, effective un…

Offline Reinforcement Learning for Rotation Profile Control in Tokamaks

2026-05-07 · Rohit Sonker, Hiro Josep Farre Kaga, Jiayu Chen, Andrew Rothstein 외 arxiv

Tokamaks remain leading candidates for achieving practical fusion energy, yet many important control problems inside these devices are still difficult or unsolved. One such challenge is controlling the plasma rotation pr…

Reinforcement LearningOffline RL