paper-with-me

홈 › Papers

Auditing Stance Asymmetry in Generative Explanations

2026-05-27 · Jiarui Han arxiv

Bias evaluation for language models has made substantial progress on bounded comparisons, such as overt derogation, stereotype association, or label-sensitive differences under controlled substitutions. Open-ended explanations raise a different problem: they guide interpretation by assigning responsibility, legitimacy, context, and grievance. A model can avoid hostile language while making one side structurally understandable and another personally at fault, overreacting, or less worth taking seriously. We call this stance-bearing asymmetry in generative explanations. We propose Symmetry Decomposition Evaluation (SDE), which tests paired situations with concrete group labels, structural-role rewrites, and explicit support or counter-evidence. In a controlled 32-family prototype suite, this decomposition shows that surface differences are not all alike: some weaken under structural or evidence control, while others remain as stable differences in how the model assigns blame, context, or legitimacy. Targeted case review and judge comparison suggest a broader difficulty for evaluating open-ended framing asymmetries: judge readings shift across operationalizations, and scalar scores can flatten distinctions that readers use to interpret explanatory stance. SDE therefore reframes generative bias evaluation as an audit of explanatory stance -- what stance each side receives, how it changes under decomposition, and where automatic scoring becomes unstable.

📄 PDF Abstract BibTeX arXiv:2605.27988

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

XAudit : A Theoretical Look at Auditing with Explanations

2022-06-09 · Chhavi Yadav, Michal Moshkovitz, Kamalika Chaudhuri

Responsible use of machine learning requires models to be audited for undesirable properties. While a body of work has proposed using explanations for auditing, how to do so and why has remained relatively ill-understood…

BIG-bench Machine LearningcounterfactualSensitivity

Auditing Local Explanations is Hard

2024-07-18 · Robi Bhattacharjee, Ulrike Von Luxburg

In sensitive contexts, providers of machine learning algorithms are increasingly required to give explanations for their algorithms' decisions. However, explanation receivers might not trust the provider, who potentially…

D4Explainer: In-Distribution GNN Explanations via Discrete Denoising Diffusion

2023-10-30 · Jialin Chen, Shirley Wu, Abhijit Gupta, Rex Ying

The widespread deployment of Graph Neural Networks (GNNs) sparks significant interest in their explainability, which plays a vital role in model auditing and ensuring trustworthy graph learning. The objective of GNN expl…

counterfactualDenoisingGraph Learning

D4Explainer: In-distribution Explanations of Graph Neural Network via Discrete Denoising Diffusion

2023-09-21 · NeurIPS 2023 11

The widespread deployment of Graph Neural Networks (GNNs) sparks significant interest in their explainability, which plays a vital role in model auditing and ensuring trustworthy graph learning. The objective of GNN expl…

Open the Black Box Data-Driven Explanation of Black Box Decision Systems

2018-06-26 · Dino Pedreschi, Fosca Giannotti, Riccardo Guidotti, Anna Monreale 외

Black box systems for automated decision making, often based on machine learning over (big) data, map a user's features into a class or a score without exposing the reasons why. This is problematic not only for lack of t…

Decision Making