paper-with-me

홈 › Papers

Attributing Fair Decisions with Attention Interventions

2021-09-08 · NAACL (TrustNLP) 2022 7 · Ninareh Mehrabi, Umang Gupta, Fred Morstatter, Greg Ver Steeg, Aram Galstyan

The widespread use of Artificial Intelligence (AI) in consequential domains, such as healthcare and parole decision-making systems, has drawn intense scrutiny on the fairness of these methods. However, ensuring fairness is often insufficient as the rationale for a contentious decision needs to be audited, understood, and defended. We propose that the attention mechanism can be used to ensure fair outcomes while simultaneously providing feature attributions to account for how a decision was made. Toward this goal, we design an attention-based model that can be leveraged as an attribution framework. It can identify features responsible for both performance and fairness of the model through attention interventions and attention weight manipulation. Using this attribution framework, we then design a post-processing bias mitigation strategy and compare it with a suite of baselines. We demonstrate the versatility of our approach by conducting experiments on two distinct data types, tabular and textual.

📄 PDF Abstract BibTeX arXiv:2109.03952

Code (1)

ninarehm/attribution 공식 구현 tf

Tasks

Decision MakingFairness

Similar Papers 제목 키워드 기반

A comparative study of fairness-enhancing interventions in machine learning

2018-02-13 · Sorelle A. Friedler, Carlos Scheidegger, Suresh Venkatasubramanian, Sonam Choudhary 외

Computers are increasingly used to make decisions that have significant impact in people's lives. Often, these predictions can affect different population subgroups disproportionately. As a result, the issue of fairness …

BIG-bench Machine LearningFairness

Fairness in Learning-Based Sequential Decision Algorithms: A Survey

2020-01-14 · Xueru Zhang, Mingyan Liu

Algorithmic fairness in decision-making has been studied extensively in static settings where one-shot decisions are made on tasks such as classification. However, in practice most decision-making processes are of a sequ…

Decision MakingFairnessSequential Decision MakingSurvey

Fairness Shields: Safeguarding against Biased Decision Makers

2024-12-16 · Filip Cano, Thomas A. Henzinger, Bettina Könighofer, Konstantin Kueffner 외

As AI-based decision-makers increasingly influence human lives, it is a growing concern that their decisions are often unfair or biased with respect to people's sensitive attributes, such as gender and race. Most existin…

Fairness

Fair outputs, Biased Internals: Causal Potency and Asymmetry of Latent Bias in LLMs for High-Stakes Decisions

2026-05-12 · Jagdish Tripathy, Marcus Buckmann arxiv

Instruction-tuned language models exhibit behavioural fairness in high-stakes decisions while retaining biased associations in their internal representations. However, whether these suppressed representations can affect …

parameter-efficient fine-tuningPrompt Engineering

Algorithmic Tradeoffs in Fair Lending: Profitability, Compliance, and Long-Term Impact

2025-05-08 · Aayam Bansal

As financial institutions increasingly rely on machine learning models to automate lending decisions, concerns about algorithmic fairness have risen. This paper explores the tradeoff between enforcing fairness constraint…

Fairness