paper-with-me

홈 › Papers

From Bias Mitigation to Bias Negotiation: Governing Identity and Sociocultural Reasoning in Generative AI

2026-02-05 · Zackary Okun Dunivin, Bingyi Han, John Bollenbocher arxiv

LLMs act in the social world by drawing upon shared cultural patterns to make social situations understandable and actionable. Because identity is often part of the inferential substrate of competent judgment, ethical alignment requires regulating when and how systems invoke identity. Yet the dominant governance regime for identity-related harm remains bias mitigation, which treats identity primarily as a source of measurable disparities or harmful associations to be detected and suppressed. This leaves underspecified a positive, context-sensitive role for identity in interpretation. We call this governance problem bias negotiation: the normative regulation of identity-conditioned judgments of sociocultural relevance, inference, and justification. Empirically, we probe the feasibility of bias negotiation through semi-structured interviews with multiple publicly deployed chatbots. We identify recurring repertoires for negotiating identity including probabilistic framing of group tendencies and harm-value balancing. We also observe failure modes in which models avoid hard tradeoffs or apply principles inconsistently. Bias negotiation matters for justice because a positive role for sociocultural reasoning is required to recognize and potentially remediate structural inequities. But it is equally implicated in core model functionality as sociocultural competence is needed for systems that operate across heterogeneous cultural contexts. Because bias negotiation is a procedural capability expressed through deliberation and interaction, it cannot be validated by static benchmarks alone. To support targeted training, we introduce a broad but explicit framework that decomposes bias negotiation into an action space of negotiation moves (what to observe and score) and a complementary set of case features (over which the model negotiates), enabling systematic test-suite design and evaluation.

📄 PDF Abstract BibTeX arXiv:2602.18459

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Tutorial on Cognitive Biases in Agentic AI-Driven 6G Autonomous Networks

2025-10-22 · Hatim Chergui, Farhad Rezazadeh, Merouane Debbah, Christos Verikoukis arxiv

The path to higher network autonomy in 6G lies beyond the mere optimization of key performance indicators (KPIs), requiring systems that perceive and reason over the network environment as it is. This can be achieved thr…

Anatomizing Bias in Facial Analysis

2021-12-13 · Richa Singh, Puspita Majumdar, Surbhi Mittal, Mayank Vatsa

Existing facial analysis systems have been shown to yield biased results against certain demographic subgroups. Due to its impact on society, it has become imperative to ensure that these systems do not discriminate base…

Bias Detection

Whither Bias Goes, I Will Go: An Integrative, Systematic Review of Algorithmic Bias Mitigation

2024-10-21 · Louis Hickman, Christopher Huynh, Jessica Gass, Brandon Booth 외

Machine learning (ML) models are increasingly used for personnel assessment and selection (e.g., resume screeners, automatically scored interviews). However, concerns have been raised throughout society that ML assessmen…

Fairness

Entropy-based Attention Regularization Frees Unintended Bias Mitigation from Lists

2022-03-17 · Findings (ACL) 2022 5 · Giuseppe Attanasio, Debora Nozza, Dirk Hovy, Elena Baralis

Natural Language Processing (NLP) models risk overfitting to specific terms in the training data, thereby reducing their performance, fairness, and generalizability. E.g., neural hate speech detection models are strongly…

Abuse DetectionBias DetectionFairnessHate Speech Detection

Investigation of Accuracy and Bias in Face Recognition Trained with Synthetic Data

2025-07-28 · Pavel Korshunov, Ketan Kotwal, Christophe Ecabert, Vidit Vidit 외 arxiv

Synthetic data has emerged as a promising alternative for training face recognition (FR) models, offering advantages in scalability, privacy compliance, and potential for bias mitigation. However, critical questions rema…

Face Recognition