paper-with-me

홈 › Papers

Which private attributes do VLMs agree on and predict well?

2026-02-08 · Olena Hrynenko, Darya Baranouskaya, Alina Elena Baia, Andrea Cavallaro arxiv

Visual Language Models (VLMs) are often used for zero-shot detection of visual attributes in the image. We present a zero-shot evaluation of open-source VLMs for privacy-related attribute recognition. We identify the attributes for which VLMs exhibit strong inter-annotator agreement, and discuss the disagreement cases of human and VLM annotations. Our results show that when evaluated against human annotations, VLMs tend to predict the presence of privacy attributes more often than human annotators. In addition to this, we find that in cases of high inter-annotator agreement between VLMs, they can complement human annotation by identifying attributes overlooked by human annotators. This highlights the potential of VLMs to support privacy annotations in large-scale image datasets.

📄 PDF Abstract BibTeX arXiv:2602.07931

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Eye of Sherlock Holmes: Uncovering User Private Attribute Profiling via Vision-Language Model Agentic Framework

2025-05-25 · Feiran Liu, Yuzhe Zhang, Xinyi Huang, Yinan Peng 외

Our research reveals a new privacy risk associated with the vision-language model (VLM) agentic framework: the ability to infer sensitive attributes (e.g., age and health information) and even abstract ones (e.g., person…

AttributeLanguage ModelingLanguage ModellingVisual Reasoning

Private Attribute Inference from Images with Vision-Language Models

2024-04-16 · Batuhan Tömekçe, Mark Vero, Robin Staab, Martin Vechev

As large language models (LLMs) become ubiquitous in our daily tasks and digital interactions, associated privacy risks are increasingly in focus. While LLM privacy research has primarily focused on the leakage of model …

Attribute

Rethinking Visual Privacy: A Compositional Privacy Risk Framework for Severity Assessment with VLMs

2026-03-23 · Efthymios Tsaprazlis, Tiantian Feng, Anil Ramakrishna, Sai Praneeth Karimireddy 외 arxiv

Existing visual privacy benchmarks largely treat privacy as a binary property, labeling images as private or non-private based on visible sensitive content. We argue that privacy is fundamentally compositional. Attribute…

Benchmarks for Vision-Language Models in Urban Perception Should Be Reliability-Aware and Negotiated

2026-05-30 · Rashid Mushkani arxiv

Vision-language models (VLMs) are increasingly used to generate structured descriptions of street-level imagery for tasks such as streetscape auditing, mapping, and public consultation. These uses combine observable attr…

Training privacy-preserving video analytics pipelines by suppressing features that reveal information about private attributes

2022-03-05 · Chau Yi Li, Andrea Cavallaro

Deep neural networks are increasingly deployed for scene analytics, including to evaluate the attention and reaction of people exposed to out-of-home advertisements. However, the features extracted by a deep neural netwo…

AttributeEmotion RecognitionPrivacy Preserving