paper-with-me

홈 › Papers

A Closed-Form Solution for Debiasing Vision-Language Models with Utility Guarantees Across Modalities and Tasks

2026-03-13 · Tangzheng Lian, Guanyu Hu, Yijing Ren, Dimitrios Kollias, Oya Celiktutan arxiv

While Vision-Language Models (VLMs) have achieved remarkable performance across diverse downstream tasks, recent studies have shown that they can inherit social biases from the training data and further propagate them into downstream applications. To address this issue, various debiasing approaches have been proposed, yet most of them aim to improve fairness without having a theoretical guarantee that the utility of the model is preserved. In this paper, we introduce a debiasing method that yields a \textbf{closed-form} solution in the cross-modal space, achieving Pareto-optimal fairness with \textbf{bounded utility losses}. Our method is \textbf{training-free}, requires \textbf{no annotated data}, and can jointly debias both visual and textual modalities across downstream tasks. Extensive experiments show that our method outperforms existing methods in debiasing VLMs across diverse fairness metrics and datasets for both group and \textbf{intersectional} fairness in downstream tasks such as zero-shot image classification, text-to-image retrieval, and text-to-image generation while preserving task performance.

📄 PDF Abstract BibTeX arXiv:2603.12998

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-Shot Image ClassificationText-to-Image GenerationImage Retrieval

Similar Papers 제목 키워드 기반

Debiasing Vision-Language Models via Biased Prompts

2023-01-31 · Ching-Yao Chuang, Varun Jampani, Yuanzhen Li, Antonio Torralba 외

Machine learning models have been shown to inherit biases from their training datasets. This can be particularly problematic for vision-language foundation models trained on uncurated datasets scraped from the internet. …

SANER: Annotation-free Societal Attribute Neutralizer for Debiasing CLIP

2024-08-19 · Yusuke Hirota, Min-Hung Chen, Chien-Yi Wang, Yuta Nakashima 외

Large-scale vision-language models, such as CLIP, are known to contain societal bias regarding protected attributes (e.g., gender, age). This paper aims to address the problems of societal bias in CLIP. Although previous…

AttributeImage GenerationText-to-Image Generation

Cognitive Debiasing Large Language Models for Decision-Making

2025-04-05 · Yougang Lyu, Shijie Ren, Yue Feng, Zihan Wang 외

Large language models (LLMs) have shown potential in supporting decision-making applications, particularly as personal conversational assistants in the financial, healthcare, and legal domains. While prompt engineering s…

Decision MakingPrompt Engineering

BendVLM: Test-Time Debiasing of Vision-Language Embeddings

2024-11-07 · Walter Gerych, Haoran Zhang, Kimia Hamidieh, Eileen Pan 외

Vision-language model (VLM) embeddings have been shown to encode biases present in their training data, such as societal biases that prescribe negative characteristics to members of various racial and gender identities. …

AttributeImage GenerationLanguage ModelingLanguage Modelling

A Prompt Array Keeps the Bias Away: Debiasing Vision-Language Models with Adversarial Learning

2022-03-22 · Hugo Berg, Siobhan Mackenzie Hall, Yash Bhalgat, Wonsuk Yang 외

Vision-language models can encode societal biases and stereotypes, but there are challenges to measuring and mitigating these multimodal harms due to lacking measurement robustness and feature degradation. To address the…