paper-with-me

홈 › Papers

Bias-Aware Machine Unlearning: Towards Fairer Vision Models via Controllable Forgetting

2025-09-09 · Sai Siddhartha Chary Aylapuram, Veeraraju Elluru, Shivang Agarwal arxiv

Deep neural networks often rely on spurious correlations in training data, leading to biased or unfair predictions in safety-critical domains such as medicine and autonomous driving. While conventional bias mitigation typically requires retraining from scratch or redesigning data pipelines, recent advances in machine unlearning provide a promising alternative for post-hoc model correction. In this work, we investigate \textit{Bias-Aware Machine Unlearning}, a paradigm that selectively removes biased samples or feature representations to mitigate diverse forms of bias in vision models. Building on privacy-preserving unlearning techniques, we evaluate various strategies including Gradient Ascent, LoRA, and Teacher-Student distillation. Through empirical analysis on three benchmark datasets, CUB-200-2011 (pose bias), CIFAR-10 (synthetic patch bias), and CelebA (gender bias in smile detection), we demonstrate that post-hoc unlearning can substantially reduce subgroup disparities, with improvements in demographic parity of up to \textbf{94.86\%} on CUB-200, \textbf{30.28\%} on CIFAR-10, and \textbf{97.37\%} on CelebA. These gains are achieved with minimal accuracy loss and with methods scoring an average of 0.62 across the 3 settings on the joint evaluation of utility, fairness, quality, and privacy. Our findings establish machine unlearning as a practical framework for enhancing fairness in deployed vision systems without necessitating full retraining.

📄 PDF Abstract BibTeX arXiv:2509.07456

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

DCAST: Diverse Class-Aware Self-Training Mitigates Selection Bias for Fairer Learning

2024-09-30 · Yasin I. Tepeli, Joana P. Gonçalves

Fairness in machine learning seeks to mitigate model bias against individuals based on sensitive features such as sex or age, often caused by an uneven representation of the population in the training data due to selecti…

DiversityDomain AdaptationFairnessMulti-class Classification+1

Unlearning Algorithmic Biases over Graphs

2025-05-20 · O. Deniz Kose, Gonzalo Mateos, Yanning Shen

The growing enforcement of the right to be forgotten regulations has propelled recent advances in certified (graph) unlearning strategies to comply with data removal requests from deployed machine learning (ML) models. M…

FairnessNode Classification

Classification-Head Bias in Class-Level Machine Unlearning: Diagnosis, Mitigation, and Evaluation

2026-05-09 · Weidong Zheng, Kongyang Chen, Yuanwei Guo, Yatie Xiao arxiv

Class-level machine unlearning aims to remove the influence of specified classes while preserving model utility on retained classes. Existing methods are commonly evaluated by retain-set accuracy, forget-set accuracy, an…

Debiasing Methods for Fairer Neural Models in Vision and Language Research: A Survey

2022-11-10 · Otávio Parraga, Martin D. More, Christian M. Oliveira, Nathan S. Gavenski 외

Despite being responsible for state-of-the-art results in several computer vision and natural language processing tasks, neural networks have faced harsh criticism due to some of their current shortcomings. One of them i…

Decision MakingFairness

Interference-Aware Multi-Task Unlearning

2026-05-18 · Ying-Hua Huang, Rui Fang, Hsi-Wen Chen, Ming-Syan Chen arxiv

Machine unlearning aims to remove the contribution of designated training data from a trained model while preserving performance on the remaining data. Existing work mainly focuses on single-task settings, whereas modern…