paper-with-me

홈 › Papers

FACET: Fairness in Computer Vision Evaluation Benchmark

2023-08-31 · ICCV 2023 1 · Laura Gustafson, Chloe Rolland, Nikhila Ravi, Quentin Duval, Aaron Adcock, Cheng-Yang Fu, Melissa Hall, Candace Ross

Computer vision models have known performance disparities across attributes such as gender and skin tone. This means during tasks such as classification and detection, model performance differs for certain classes based on the demographics of the people in the image. These disparities have been shown to exist, but until now there has not been a unified approach to measure these differences for common use-cases of computer vision models. We present a new benchmark named FACET (FAirness in Computer Vision EvaluaTion), a large, publicly available evaluation set of 32k images for some of the most common vision tasks - image classification, object detection and segmentation. For every image in FACET, we hired expert reviewers to manually annotate person-related attributes such as perceived skin tone and hair type, manually draw bounding boxes and label fine-grained person-related classes such as disk jockey or guitarist. In addition, we use FACET to benchmark state-of-the-art vision models and present a deeper understanding of potential performance disparities and challenges across sensitive demographic attributes. With the exhaustive annotations collected, we probe models using single demographics attributes as well as multiple attributes using an intersectional approach (e.g. hair color and perceived skin tone). Our results show that classification, detection, segmentation, and visual grounding models exhibit performance disparities across demographic attributes and intersections of attributes. These harms suggest that not all people represented in datasets receive fair and equitable treatment in these vision tasks. We hope current and future results using our benchmark will contribute to fairer, more robust vision models. FACET is available publicly at https://facet.metademolab.com/

📄 PDF Abstract BibTeX arXiv:2309.00035

Code (0)

등록된 구현이 없습니다.

Tasks

Fairnessimage-classificationImage Classificationobject-detectionObject DetectionVisual Grounding

Similar Papers 제목 키워드 기반

Evaluating Fairness in Large Vision-Language Models Across Diverse Demographic Attributes and Prompts

2024-06-25 · Xuyang Wu, YuAn Wang, Hsin-Tai Wu, Zhiqiang Tao 외

Large vision-language models (LVLMs) have recently achieved significant progress, demonstrating strong capabilities in open-world visual understanding. However, it is not yet clear how LVLMs address demographic biases in…

FairnessQuestion AnsweringSingle Choice QuestionVisual Question Answering

mFARM: Towards Multi-Faceted Fairness Assessment based on HARMs in Clinical Decision Support

2025-09-02 · Shreyash Adappanavar, Krithi Shailya, Gokul S Krishnan, Sriraam Natarajan 외 arxiv

The deployment of Large Language Models (LLMs) in high-stakes medical settings poses a critical AI alignment challenge, as models can inherit and amplify societal biases, leading to significant disparities. Existing fair…

Ethical AI on the Waitlist: Group Fairness Evaluation of LLM-Aided Organ Allocation

2025-03-29 · Hannah Murray, Brian Hyeongseok Kim, Isabelle Lee, Jason Byun 외

Large Language Models (LLMs) are becoming ubiquitous, promising automation even in high-stakes scenarios. However, existing evaluation methods often fall short -- benchmarks saturate, accuracy-based metrics are overly si…

Fairness

Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models

2024-05-24 · Byung-Kwan Lee, Chae Won Kim, Beomchan Park, Yong Man Ro

The rapid development of large language and vision models (LLVMs) has been driven by advances in visual instruction tuning. Recently, open-source LLVMs have curated high-quality visual instruction tuning datasets and uti…

Common Sense ReasoningLanguage ModellingMambaMath+2

Federated Learning at the Forefront of Fairness: A Multifaceted Perspective

2026-01-31 · Noorain Mukhtiar, Adnan Mahmood, Yipeng Zhou, Jian Yang 외 arxiv

Fairness in Federated Learning (FL) is emerging as a critical factor driven by heterogeneous clients' constraints and balanced model performance across various scenarios. In this survey, we delineate a comprehensive clas…

Federated Learning