paper-with-me

홈 › Papers

Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment

2025-09-24 · Aravind Narayanan, Vahid Reza Khazaie, Shaina Raza arxiv

Large vision-language models (VLMs) can jointly interpret images and text, but they are also prone to absorbing and reproducing harmful social stereotypes when visual cues such as age, gender, race, clothing, or occupation are present. To investigate these risks, we introduce a news-image benchmark consisting of 1,343 image-question pairs drawn from diverse outlets, which we annotated with ground-truth answers and demographic attributes (age, gender, race, occupation, and sports). We evaluate a range of state-of-the-art VLMs and employ a large language model (LLM) as judge, with human verification. Our findings show that: (i) visual context systematically shifts model outputs in open-ended settings; (ii) bias prevalence varies across attributes and models, with particularly high risk for gender and occupation; and (iii) higher faithfulness does not necessarily correspond to lower bias. We release the benchmark prompts, evaluation rubric, and code to support reproducible and fairness-aware multimodal assessment.

📄 PDF Abstract BibTeX arXiv:2509.19659

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias

2024-10-17 · Kuleen Sasse, Shan Chen, Jackson Pond, Danielle Bitterman 외

As Vision Language Models (VLMs) gain widespread use, their fairness remains under-explored. In this paper, we analyze demographic biases across five models and six datasets. We find that portrait datasets like UTKFace a…

BenchmarkingBias DetectionFairnessLanguage Modeling+1

ViLBias: A Comprehensive Framework for Bias Detection through Linguistic and Visual Cues , presenting Annotation Strategies, Evaluation, and Key Challenges

2024-12-22 · Shaina Raza, Caesar Saleh, Emrul Hasan, Franklin Ogidi 외

The integration of Large Language Models (LLMs) and Vision-Language Models (VLMs) opens new avenues for addressing complex challenges in multimodal content analysis, particularly in biased news detection. This study intr…

Bias Detection

My Answer Is NOT 'Fair': Mitigating Social Bias in Vision-Language Models via Fair and Biased Residuals

2025-05-26 · Jian Lan, Yifei Fu, Udo Schlegel, Gengyuan Zhang 외

Social bias is a critical issue in large vision-language models (VLMs), where fairness- and ethics-related problems harm certain groups of people in society. It is unknown to what extent VLMs yield social bias in generat…

EthicsFairnessMultiple-choice

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

2025-05-28 · Chahat Raj, Bowen Wei, Aylin Caliskan, Antonios Anastasopoulos 외

While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less attention. Existing VLM bias studies often focus on portrait-style images and…

Decision MakingQuestion AnsweringVisual Question Answering (VQA)

Multi-Modal Framing Analysis of News

2025-03-26 · Arnav Arora, Srishti Yadav, Maria Antoniak, Serge Belongie 외

Automated frame analysis of political communication is a popular task in computational social science that is used to study how authors select aspects of a topic to frame its reception. So far, such studies have been nar…