Towards Auditability for Fairness in Deep Learning
Group fairness metrics can detect when a deep learning model behaves differently for advantaged and disadvantaged groups, but even models that score well on these metrics can make blatantly unfair predictions. We present smooth prediction sensitivity, an efficiently computed measure of individual fairness for deep learning models that is inspired by ideas from interpretability in deep learning. smooth prediction sensitivity allows individual predictions to be audited for fairness. We present preliminary experimental results suggesting that smooth prediction sensitivity can help distinguish between fair and unfair predictions, and that it may be helpful in detecting blatantly unfair predictions from "group-fair" models.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep LearningFairnessPredictionSensitivityMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Auditability and the Landscape of Distance to Multicalibration
Calibration is a critical property for establishing the trustworthiness of predictors that provide uncertainty estimates. Multicalibration is a strengthening of calibration which requires that predictors be calibrated on…
The Explabox: Model-Agnostic Machine Learning Transparency & Analysis
We present the Explabox: an open-source toolkit for transparent and responsible machine learning (ML) model development and usage. Explabox aids in achieving explainable, fair and robust models by employing a four-step s…
DescriptiveFairnessAssessing the Auditability of AI-integrating Systems: A Framework and Learning Analytics Case Study
Audits contribute to the trustworthiness of Learning Analytics (LA) systems that integrate Artificial Intelligence (AI) and may be legally required in the future. We argue that the efficacy of an audit depends on the aud…
EthicsFortifying Federated Learning Towards Trustworthiness via Auditable Data Valuation and Verifiable Client Contribution
Ensuring auditability and verifiability in FL is both challenging and essential to guarantee that local data remains untampered and client updates are trustworthy. Recent FL frameworks assess client contributions thr…
Data PoisoningData ValuationFairnessFederated Learning+1Enactive Artificial Intelligence: Subverting Gender Norms in Robot-Human Interaction
This paper introduces Enactive Artificial Intelligence (eAI) as an intersectional gender-inclusive stance towards AI. AI design is an enacted human sociocultural practice that reflects human culture and values. Unreprese…
Cultural Vocal Bursts Intensity PredictionEthicsFairness