paper-with-me

홈 › Papers

Perception of Visual Content: Differences Between Humans and Foundation Models

2024-11-28 · Nardiena A. Pratama, Shaoyang Fan, Gianluca Demartini

Human-annotated content is often used to train machine learning (ML) models. However, recently, language and multi-modal foundational models have been used to replace and scale-up human annotator's efforts. This study compares human-generated and ML-generated annotations of images representing diverse socio-economic contexts. We aim to understand differences in perception and identify potential biases in content interpretation. Our dataset comprises images of people from various geographical regions and income levels, covering various daily activities and home environments. We compare human and ML-generated annotations semantically and evaluate their impact on predictive models. Our results show highest similarity between ML captions and human labels from a low-level perspective, i.e., types of words that appear and sentence structures, but all three annotations are alike in how similar or dissimilar they perceive images across different regions. Additionally, ML Captions resulted in best overall region classification performance, while ML Objects and ML Captions performed best overall for income regression. The varying performance of annotation sets highlights the notion that all annotations are important, and that human-generated annotations are yet to be replaceable.

📄 PDF Abstract BibTeX arXiv:2411.18968

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Similar Papers 제목 키워드 기반

MVP-Bench: Can Large Vision--Language Models Conduct Multi-level Visual Perception Like Humans?

2024-10-06 · Guanzhen Li, Yuxi Xie, Min-Yen Kan

Humans perform visual perception at multiple levels, including low-level object recognition and high-level semantic interpretation such as behavior understanding. Subtle differences in low-level details can lead to subst…

Object Recognition

Diptychs of human and machine perceptions

2020-10-12 · Vivien Cabannes, Thomas Kerdreux, Louis Thiry

We propose visual creations that put differences in algorithms and humans \emph{perceptions} into perspective. We exploit saliency maps of neural networks and visual focus of humans to create diptychs that are reinterpre…

Do Computational Models Differ Systematically From Human Object Perception?

2016-06-01 · CVPR 2016 6 · R. T. Pramod, S. P. Arun

Recent advances in neural networks have revolutionized computer vision, but these algorithms are still outperformed by humans. Could this performance gap be due to systematic differences between object representations in…

Object

Comparative Analysis Of Color Models For Human Perception And Visual Color Difference

2024-06-27 · Aruzhan Burambekova, Pakizar Shamoi

Color is integral to human experience, influencing emotions, decisions, and perceptions. This paper presents a comparative analysis of various color models' alignment with human visual perception. The study evaluates col…

Models Alignment

Understanding Why ChatGPT Outperforms Humans in Visualization Design Advice

2025-08-03 · Yongsu Ahn, Nam Wook Kim arxiv

This paper investigates why recent generative AI models outperform humans in data visualization knowledge tasks. Through systematic comparative analysis of responses to visualization questions, we find that differences e…