paper-with-me

홈 › Papers

ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models

2025-08-02 · Chuangchuang Tan, Jinglu Wang, Xiang Ming, Renshuai Tao, Yunchao Wei, Yao Zhao, Yan Lu arxiv

Advances in generative models have led to AI-generated images visually indistinguishable from authentic ones. Despite numerous studies on detecting AI-generated images with classifiers, a gap persists between such methods and human cognitive forensic analysis. We present ForenX, a novel method that not only identifies the authenticity of images but also provides explanations that resonate with human thoughts. ForenX employs the powerful multimodal large language models (MLLMs) to analyze and interpret forensic cues. Furthermore, we overcome the limitations of standard MLLMs in detecting forgeries by incorporating a specialized forensic prompt that directs the MLLMs attention to forgery-indicative attributes. This approach not only enhance the generalization of forgery detection but also empowers the MLLMs to provide explanations that are accurate, relevant, and comprehensive. Additionally, we introduce ForgReason, a dataset dedicated to descriptions of forgery evidences in AI-generated images. Curated through collaboration between an LLM-based agent and a team of human annotators, this process provides refined data that further enhances our model's performance. We demonstrate that even limited manual annotations significantly improve explanation quality. We evaluate the effectiveness of ForenX on two major benchmarks. The model's explainability is verified by comprehensive subjective evaluations.

📄 PDF Abstract BibTeX arXiv:2508.01402

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

2025-11-28 · Huangsen Cao, Qin Mei, Zhiheng Li, Yuxi Li 외 arxiv

The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing growing risks to social trust and information integrity. This motivates detect…

Reinforcement LearningDomain Generalization

FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models

2024-04-20 · Yixuan Li, Xuelin Liu, Xiaoyang Wang, Bu Sung Lee 외

The ability to distinguish whether an image is generated by artificial intelligence (AI) is a crucial ingredient in human intelligence, usually accompanied by a complex and dialectical forensic and reasoning process. How…

Binary ClassificationFake Image DetectionQuestion Answering

BusterX: MLLM-Powered AI-Generated Video Forgery Detection and Explanation

2025-05-19 · Haiquan Wen, Yiwei He, Zhenglin Huang, Tianxiao Li 외

Advances in AI generative models facilitate super-realistic video synthesis, amplifying misinformation risks via social media and eroding trust in digital content. Several research works have explored new deepfake detect…

Binary ClassificationDeepFake DetectionFace SwappingLarge Language Model+3

SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs

2025-11-16 · Shail Desai, Aditya Pawar, Li Lin, Xin Wang 외 arxiv

Artificial Intelligence (AI) has made it possible for anyone to create images, audio, and video with unprecedented ease, enriching education, communication, and creative expression. At the same time, the rapid rise of AI…

DeepFake Detection

Explainable AI-Generated Image Detection RewardBench

2025-11-15 · Michael Yang, Shijian Deng, William T. Doan, Kai Wang 외 arxiv

Conventional, classification-based AI-generated image detection methods cannot explain why an image is considered real or AI-generated in a way a human expert would, which reduces the trustworthiness and persuasiveness o…

Image Generation