paper-with-me

홈 › Papers

A Framework for Generating Semantically Ambiguous Images to Probe Human and Machine Perception

2026-03-25 · Yuqi Hu, Vasha DuTell, Ahna R. Girshick, Jennifer E. Corbett arxiv

The classic duck-rabbit illusion reveals that when visual evidence is ambiguous, the human brain must decide what it sees. But where exactly do human observers draw the line between ''duck'' and ''rabbit'', and do machine classifiers draw it in the same place? We use semantically ambiguous images as interpretability probes to expose how vision models represent the boundaries between concepts. We present a psychophysically-informed framework that interpolates between concepts in the CLIP embedding space to generate continuous spectra of ambiguous images, allowing us to precisely measure where and how humans and machine classifiers place their semantic boundaries. Using this framework, we show that machine classifiers are more biased towards seeing ''rabbit'', whereas humans are more aligned with the CLIP embedding used for synthesis, and the guidance scale seems to affect human sensitivity more strongly than machine classifiers. Our framework demonstrates how controlled ambiguity can serve as a diagnostic tool to bridge the gap between human psychophysical analysis, image classification, and generative image models, offering insight into human-model alignment, robustness, model interpretability, and image synthesis methods.

📄 PDF Abstract BibTeX arXiv:2603.24730

Code (0)

등록된 구현이 없습니다.

Tasks

Image Classification

Similar Papers 제목 키워드 기반

Can large language models generate salient negative statements?

2023-05-26 · Hiba Arnaout, Simon Razniewski

We examine the ability of large language models (LLMs) to generate salient (interesting) negative statements about real-world entities; an emerging research topic of the last few years. We probe the LLMs using zero- and …

Negation

Underspecification in Scene Description-to-Depiction Tasks

2022-10-11 · Ben Hutchinson, Jason Baldridge, Vinodkumar Prabhakaran

Questions regarding implicitness, ambiguity and underspecification are crucial for understanding the task validity and ethical concerns of multimodal image+text systems, yet have received little attention to date. This p…

Position

When Visual Evidence is Ambiguous: Pareidolia as a Diagnostic Probe for Vision Models

2026-03-04 · Qianpu Chen, Derya Soydaner, Rob Saunders arxiv

When visual evidence is ambiguous, vision models must decide how to interpret face-like patterns. Face pareidolia, the perception of faces in non-face objects, provides a controlled probe of such decisions. We introduce …

Object DetectionFace Detection

Semantic Robustness Probing via Inpainting: An Interactive Tool for Safety-Critical Object Detection

2026-05-26 · Nico Steckhan, Krutarth Prajapati, Weija Shao, Silvia Vock arxiv

Testing object detectors in safety-critical domains requires semantically meaningful probes beyond pixel-level corruptions. We present SemProbe, a tool for semantic robustness probing: users upload deployment images, cre…

Object Detection

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

2026-06-03 · Govind Ramesh, Yao Dou, Wei Xu arxiv

Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, while existing attribution methods are designed to explain observable o…