paper-with-me

홈 › Papers

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals

2026-05-21 · Zihang Lin, Huaiyuan Qin, Muli Yang, Hongyuan Zhu arxiv

Assessing progress toward the Sustainable Development Goals (SDGs) requires multi-step reasoning over visual cues, contextual knowledge, and development indicators, where incomplete evidence use and imperfect evidence integration can introduce hidden prediction biases. Real-world SDG monitoring further spans both qualitative judgments and quantitative estimation. However, existing benchmarks typically evaluate these aspects in isolation, obscuring systematic biases that emerge when models substitute priors for evidence. To address this gap, we propose SDGBiasBench, a large-scale benchmark suite for SDG-oriented vision-language reasoning. Spanning 500k expert-involved multiple-choice questions and 50k regression tasks, the benchmark enables comprehensive assessment of both decision-level and estimation-level bias in Vision--Language Models (VLMs). Evaluations on SDGBiasBench reveal an intrinsic SDG bias in current VLMs, where predictions are frequently driven by SDG specific priors rather than reliable multi-modal cues. To mitigate such bias, we propose CADE (Contrastive Adaptive Debias Ensemble), a training-free, plug-and-play method that leverages modality-specific answer priors. CADE yields significant gains on the proposed benchmark, improving multiple-choice accuracy by up to 25% and reducing regression MAE by up to 12 points across multiple VLMs. We hope our work can foster the development of more fair and reliable AI systems for sustainable development.

📄 PDF Abstract BibTeX arXiv:2605.21919

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Benchmarking and Mitigating MCQA Selection Bias of Large Vision-Language Models

2025-09-20 · Md. Atabuzzaman, Ali Asgarov, Chris Thomas arxiv

Large Vision-Language Models (LVLMs) have achieved strong performance on vision-language tasks, particularly Visual Question Answering (VQA). While prior work has explored unimodal biases in VQA, the problem of selection…

Visual Question AnsweringSemantic SimilarityVisual Reasoning

Measuring Social Biases in Grounded Vision and Language Embeddings

2020-02-20 · NAACL 2021 4 · Candace Ross, Boris Katz, Andrei Barbu

We generalize the notion of social biases from language embeddings to grounded vision and language embeddings. Biases are present in grounded embeddings, and indeed seem to be equally or more significant than for ungroun…

Word Embeddings

Discovering and Mitigating Visual Biases through Keyword Explanation

2023-01-26 · CVPR 2024 1 · Younghyun Kim, Sangwoo Mo, Minkyu Kim, Kyungmin Lee 외

Addressing biases in computer vision models is crucial for real-world AI deployments. However, mitigating visual biases is challenging due to their unexplainable nature, often identified indirectly through visualization …

Image ClassificationImage Generation

debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias

2024-10-17 · Kuleen Sasse, Shan Chen, Jackson Pond, Danielle Bitterman 외

As Vision Language Models (VLMs) gain widespread use, their fairness remains under-explored. In this paper, we analyze demographic biases across five models and six datasets. We find that portrait datasets like UTKFace a…

BenchmarkingBias DetectionFairnessLanguage Modeling+1

White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs

2024-04-16 · Yixin Wan, Kai-Wei Chang

Social biases can manifest in language agency. However, very limited research has investigated such biases in Large Language Model (LLM)-generated content. In addition, previous works often rely on string-matching techni…

BenchmarkingLanguage ModellingLarge Language ModelSentence+1