paper-with-me

홈 › Papers

UnStar: Unlearning with Self-Taught Anti-Sample Reasoning for LLMs

2024-10-22 · Yash Sinha, Murari Mandal, Mohan Kankanhalli

The key components of machine learning are data samples for training, model for learning patterns, and loss function for optimizing accuracy. Analogously, unlearning can potentially be achieved through anti-data samples (or anti-samples), unlearning method, and reversed loss function. While prior research has explored unlearning methods and reversed loss functions, the potential of anti-samples remains largely untapped. In this paper, we introduce UnSTAR: Unlearning with Self-Taught Anti-Sample Reasoning for large language models (LLMs). Our contributions are threefold; first, we propose a novel concept of anti-sample-induced unlearning; second, we generate anti-samples by leveraging misleading rationales, which help reverse learned associations and accelerate the unlearning process; and third, we enable fine-grained targeted unlearning, allowing for the selective removal of specific associations without impacting related knowledge - something not achievable by previous works. Results demonstrate that anti-samples offer an efficient, targeted unlearning strategy for LLMs, opening new avenues for privacy-preserving machine learning and model modification.

📄 PDF Abstract BibTeX arXiv:2410.17050

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy Preserving

Similar Papers 제목 키워드 기반

Deep Self-Taught Learning for Weakly Supervised Object Localization

2017-04-18 · CVPR 2017 7 · Zequn Jie, Yunchao Wei, Xiaojie Jin, Jiashi Feng 외

Most existing weakly supervised localization (WSL) approaches learn detectors by finding positive bounding boxes based on features learned with image-level supervision. However, those features do not contain spatial loca…

ObjectObject LocalizationWeakly Supervised Object DetectionWeakly-Supervised Object Localization

Autoencoder Based Sample Selection for Self-Taught Learning

2018-08-05 · Siwei Feng, Han Yu, Marco F. Duarte

Self-taught learning is a technique that uses a large number of unlabeled data as source samples to improve the task performance on target samples. Compared with other transfer learning techniques, self-taught learning c…

Transfer Learning

Approximate Machine Unlearning through Manifold Representation Forgetting Guided by Self Mode Connectivity

2026-05-20 · Weiqi Wang, Zhiyi Tian, Chenhan Zhang, Luoyu Chen 외 arxiv

Machine unlearning is a fundamental mechanism that enforces the right to be forgotten. Existing unlearning studies that rely on label manipulation or task-gradient reversal often deliver limited unlearning effectiveness.…

Semantic Similarity

SMS: Self-supervised Model Seeding for Verification of Machine Unlearning

2025-09-30 · Weiqi Wang, Chenhan Zhang, Zhiyi Tian, Shui Yu arxiv

Many machine unlearning methods have been proposed recently to uphold users' right to be forgotten. However, offering users verification of their data removal post-unlearning is an important yet under-explored problem. C…

Hypersonic Flow Control: Generalized Deep Reinforcement Learning for Hypersonic Intake Unstart Control under Uncertainty

2026-01-27 · Trishit Mondal, Ameya D. Jagtap arxiv

The hypersonic unstart phenomenon poses a major challenge to reliable air-breathing propulsion at Mach 5 and above, where strong shock-boundary-layer interactions and rapid pressure fluctuations can destabilize inlet ope…

Zero-shot GeneralizationReinforcement Learning