paper-with-me

Papers

FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models

2024-04-20 · Yixuan Li, Xuelin Liu, Xiaoyang Wang, Bu Sung Lee, Shiqi Wang, Anderson Rocha, Weisi Lin

The ability to distinguish whether an image is generated by artificial intelligence (AI) is a crucial ingredient in human intelligence, usually accompanied by a complex and dialectical forensic and reasoning process. However, current fake image detection models and databases focus on binary classification without understandable explanations for the general populace. This weakens the credibility of authenticity judgment and may conceal potential model biases. Meanwhile, large multimodal models (LMMs) have exhibited immense visual-text capabilities on various tasks, bringing the potential for explainable fake image detection. Therefore, we pioneer the probe of LMMs for explainable fake image detection by presenting a multimodal database encompassing textual authenticity descriptions, the FakeBench. For construction, we first introduce a fine-grained taxonomy of generative visual forgery concerning human perception, based on which we collect forgery descriptions in human natural language with a human-in-the-loop strategy. FakeBench examines LMMs with four evaluation criteria: detection, reasoning, interpretation and fine-grained forgery analysis, to obtain deeper insights into image authenticity-relevant capabilities. Experiments on various LMMs confirm their merits and demerits in different aspects of fake image detection tasks. This research presents a paradigm shift towards transparency for the fake image detection area and reveals the need for greater emphasis on forensic elements in visual-language research and AI risk control. FakeBench will be available at https://github.com/Yixuan423/FakeBench.

📄 PDF Abstract BibTeX arXiv:2404.13306

Code (1)

yixuan423/fakebench 공식 구현 pytorch

Tasks

Binary ClassificationFake Image DetectionQuestion Answering

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

DeepfakeBench: A Comprehensive Benchmark of Deepfake Detection

2023-07-04 · NeurIPS 2023 11 · Zhiyuan Yan, Yong Zhang, Xinhang Yuan, Siwei Lyu 외

A critical yet frequently overlooked challenge in the field of deepfake detection is the lack of a standardized, unified, comprehensive benchmark. This issue leads to unfair performance comparisons and potentially mislea…

DeepFake DetectionFace Swapping

AVFakeBench: A Comprehensive Audio-Video Forgery Detection Benchmark for AV-LMMs

2025-11-26 · Shuhan Xia, Peipei Li, Xuannan Liu, Dongsen Zhang 외 arxiv

The threat of Audio-Video (AV) forgery is rapidly evolving beyond human-centric deepfakes to include more diverse manipulations across complex natural scenes. However, existing benchmarks are still confined to DeepFake-b…

From Manipulation to Mistrust: Explaining Diverse Micro-Video Misinformation for Robust Debunking in the Wild

2026-03-26 · Zhi Zeng, Yifei Yang, Jiaying Wu, Xulang Zhang 외 arxiv

The rise of micro-videos has reshaped how misinformation spreads, amplifying its speed, reach, and impact on public trust. Existing benchmarks typically focus on a single deception type, overlooking the diversity of real…

DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection

2025-10-26 · Kangran Zhao, Yupeng Chen, Xiaoyu Zhang, Yize Chen 외 arxiv

The misuse of advanced generative AI models has resulted in the widespread proliferation of falsified data, particularly forged human-centric audiovisual content, which poses substantial societal risks (e.g., financial f…

DeepFake Detection

MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs

2024-06-13 · Xuannan Liu, Zekun Li, Peipei Li, Shuhan Xia 외

Current multimodal misinformation detection (MMD) methods often assume a single source and type of forgery for each sample, which is insufficient for real-world scenarios where multiple forgery sources coexist. The lack …

Misinformation