paper-with-me

홈 › Papers

BriarPatches: Pixel-Space Interventions for Inducing Demographic Parity

2018-12-17 · Alexey A. Gritsenko, Alex D'Amour, James Atwood, Yoni Halpern, D. Sculley

We introduce the BriarPatch, a pixel-space intervention that obscures sensitive attributes from representations encoded in pre-trained classifiers. The patches encourage internal model representations not to encode sensitive information, which has the effect of pushing downstream predictors towards exhibiting demographic parity with respect to the sensitive information. The net result is that these BriarPatches provide an intervention mechanism available at user level, and complements prior research on fair representations that were previously only applicable by model developers and ML experts.

📄 PDF Abstract BibTeX arXiv:1812.06869

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Controllable Biases in Language Generation

2020-05-01 · Findings of the Association for Computational Linguistics 2020 · Emily Sheng, Kai-Wei Chang, Premkumar Natarajan, Nanyun Peng

We present a general approach towards controllable societal biases in natural language generation (NLG). Building upon the idea of adversarial triggers, we develop a method to induce societal biases in generated text whe…

Dialogue GenerationText Generation

Mitigating Bias with Words: Inducing Demographic Ambiguity in Face Recognition Templates by Text Encoding

2025-12-05 · Tahar Chettaoui, Naser Damer, Fadi Boutros arxiv

Face recognition (FR) systems are often prone to demographic biases, partially due to the entanglement of demographic-specific information with identity-relevant features in facial embeddings. This bias is extremely crit…

Face VerificationFace Recognition

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

2026-02-02 · Raphaël Sarfati, Eric Bigelow, Daniel Wurgaft, Siddharth Boppana 외 arxiv

Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these beliefs are encoded in representation space, how they update with new evi…

Intervention Design for Effective Sim2Real Transfer

2020-12-03 · Melissa Mozifian, Amy Zhang, Joelle Pineau, David Meger

The goal of this work is to address the recent success of domain randomization and data augmentation for the sim2real setting. We explain this success through the lens of causal inference, positioning domain randomizatio…

Causal InferenceData Augmentation

A New Paradigm for Counterfactual Reasoning in Fairness and Recourse

2024-01-25 · Lucius E. J. Bynum, Joshua R. Loftus, Julia Stoyanovich

Counterfactuals and counterfactual reasoning underpin numerous techniques for auditing and understanding artificial intelligence (AI) systems. The traditional paradigm for counterfactual reasoning in this literature is t…

counterfactualCounterfactual ReasoningFairness