paper-with-me

홈 › Papers

LLMs Are Not Yet Ready for Deepfake Image Detection

2025-06-12 · Shahroz Tariq, David Nguyen, M. A. P. Chamikara, Tingmin Wu, Alsharif Abuadbba, Kristen Moore

The growing sophistication of deepfakes presents substantial challenges to the integrity of media and the preservation of public trust. Concurrently, vision-language models (VLMs), large language models enhanced with visual reasoning capabilities, have emerged as promising tools across various domains, sparking interest in their applicability to deepfake detection. This study conducts a structured zero-shot evaluation of four prominent VLMs: ChatGPT, Claude, Gemini, and Grok, focusing on three primary deepfake types: faceswap, reenactment, and synthetic generation. Leveraging a meticulously assembled benchmark comprising authentic and manipulated images from diverse sources, we evaluate each model's classification accuracy and reasoning depth. Our analysis indicates that while VLMs can produce coherent explanations and detect surface-level anomalies, they are not yet dependable as standalone detection systems. We highlight critical failure modes, such as an overemphasis on stylistic elements and vulnerability to misleading visual patterns like vintage aesthetics. Nevertheless, VLMs exhibit strengths in interpretability and contextual analysis, suggesting their potential to augment human expertise in forensic workflows. These insights imply that although general-purpose models currently lack the reliability needed for autonomous deepfake detection, they hold promise as integral components in hybrid or human-in-the-loop detection frameworks.

📄 PDF Abstract BibTeX arXiv:2506.10474

Code (0)

등록된 구현이 없습니다.

Tasks

DeepFake DetectionFace SwappingVisual Reasoning

Similar Papers 제목 키워드 기반

Can Multi-modal (reasoning) LLMs work as deepfake detectors?

2025-03-25 · Simiao Ren, Yao Yao, Kidus Zewde, Zisheng Liang 외

Deepfake detection remains a critical challenge in the era of advanced generative models, particularly as synthetic media becomes more sophisticated. In this study, we explore the potential of state of the art multi-moda…

DeepFake DetectionFace Swapping

Investigating the Viability of Employing Multi-modal Large Language Models in the Context of Audio Deepfake Detection

2026-01-02 · Akanksha Chuchra, Shukesh Reddy, Sudeepta Mishra, Abhijit Das 외 arxiv

While Vision-Language Models (VLMs) and Multimodal Large Language Models (MLLMs) have shown strong generalisation in detecting image and video deepfakes, their use for audio deepfake detection remains largely unexplored.…

Audio Deepfake Detection

Deepfake Video Forensics based on Transfer Learning

2020-04-29 · Rahul U, Ragul M, Raja Vignesh K, Tejeswinee K

Deeplearning has been used to solve complex problems in various domains. As it advances, it also creates applications which become a major threat to our privacy, security and even to our Democracy. Such an application wh…

DeepFake DetectionFace Swappingimage-classificationImage Classification+2

Addressing Deepfake Issue in Selfie banking through camera based authentication

2025-08-27 · Subhrojyoti Mukherjee, Manoranjan Mohanty arxiv

Fake images in selfie banking are increasingly becoming a threat. Previously, it was just Photoshop, but now deep learning technologies enable us to create highly realistic fake identities, which fraudsters exploit to by…

Camera LocalizationDeepFake Detection

Scaling Laws for Deepfake Detection

2025-10-18 · Wenhao Wang, Longqi Cai, Taihong Xiao, Yuxiao Wang 외 arxiv

This paper presents a systematic study of scaling laws for the deepfake detection task. Specifically, we analyze the model performance against the number of real image domains, deepfake generation methods, and training i…

DeepFake Detection