paper-with-me

홈 › Papers

Seeing is not always believing: Benchmarking Human and Model Perception of AI-Generated Images

2023-09-26 · NeurIPS 2023 11

Photos serve as a way for humans to record what they experience in their daily lives, and they are often regarded as trustworthy sources of information. However, there is a growing concern that the advancement of artificial intelligence (AI) technology may produce fake photos, which can create confusion and diminish trust in photographs. This study aims to comprehensively evaluate agents for distinguishing state-of-the-art AI-generated visual content. Our study benchmarks both human capability and cutting-edge fake image detection AI algorithms, using a newly collected large-scale fake image dataset Fake2M. In our human perception evaluation, titled HPBench, we discovered that humans struggle significantly to distinguish real photos from AI-generated ones, with a misclassification rate of 38.7\%. Along with this, we conduct the model capability of AI-Generated images detection evaluation MPBench and the top-performing model from MPBench achieves a 13\% failure rate under the same setting used in the human evaluation. We hope that our study can raise awareness of the potential risks of AI-generated images and facilitate further research to prevent the spread of false information. More information can refer to https://github.com/Inf-imagine/Sentry.

📄 PDF Abstract BibTeX

Code (1)

inf-imagine/sentry 공식 구현

Similar Papers 제목 키워드 기반

Seeing through the Brain: Image Reconstruction of Visual Perception from Human Brain Signals

2023-07-27 · Yu-Ting Lan, Kan Ren, Yansen Wang, Wei-Long Zheng 외

Seeing is believing, however, the underlying mechanism of how human visual perceptions are intertwined with our cognitions is still a mystery. Thanks to the recent advances in both neuroscience and artificial intelligenc…

EEGImage ReconstructionTime Series

Seeing but Not Believing: Probing the Disconnect Between Visual Attention and Answer Correctness in VLMs

2025-10-20 · Zhining Liu, Ziyi Chen, Hui Liu, Chen Luo 외 arxiv

Vision-Language Models (VLMs) achieve strong results on multimodal tasks such as visual question answering, yet they can still fail even when the correct visual evidence is present. In this work, we systematically invest…

Visual Question Answering

Is Seeing Believing? Evaluating Human Sensitivity to Synthetic Video

2026-03-14 · David Wegmann, Emil Stevnsborg, Søren Knudsen, Luca Rossi 외 arxiv

Advances in machine learning have enabled the creation of realistic synthetic videos known as deepfakes. As deepfakes proliferate, concerns about rapid spread of disinformation and manipulation of public perception are m…

Seeing Is Believing? A Benchmark for Multimodal Large Language Models on Visual Illusions and Anomalies

2026-02-02 · Wenjin Hou, Wei Liu, Han Hu, Xiaoxiao Sun 외 arxiv

Multimodal Large Language Models (MLLMs) have shown remarkable proficiency on general-purpose vision-language benchmarks, reaching or even exceeding human-level performance. However, these evaluations typically rely on s…

Visual Reasoning

Seeing is not Believing: An Identity Hider for Human Vision Privacy Protection

2023-07-02 · Tao Wang, Yushu Zhang, Zixuan Yang, Xiangli Xiao 외

Massive captured face images are stored in the database for the identification of individuals. However, these images can be observed unintentionally by data managers, which is not at the will of individuals and may cause…

Appearance TransferAttributeDisentanglementFace Generation