paper-with-me

홈 › Papers

Seeing Through Deepfakes: A Human-Inspired Framework for Multi-Face Detection

2025-07-20 · Juan Hu, Shaojing Fan, Terence Sim arxiv

Multi-face deepfake videos are becoming increasingly prevalent, often appearing in natural social settings that challenge existing detection methods. Most current approaches excel at single-face detection but struggle in multi-face scenarios, due to a lack of awareness of crucial contextual cues. In this work, we develop a novel approach that leverages human cognition to analyze and defend against multi-face deepfake videos. Through a series of human studies, we systematically examine how people detect deepfake faces in social settings. Our quantitative analysis reveals four key cues humans rely on: scene-motion coherence, inter-face appearance compatibility, interpersonal gaze alignment, and face-body consistency. Guided by these insights, we introduce \textsf{HICOM}, a novel framework designed to detect every fake face in multi-face scenarios. Extensive experiments on benchmark datasets show that \textsf{HICOM} improves average accuracy by 3.3\% in in-dataset detection and 2.8\% under real-world perturbations. Moreover, it outperforms existing methods by 5.8\% on unseen datasets, demonstrating the generalization of human-inspired cues. \textsf{HICOM} further enhances interpretability by incorporating an LLM to provide human-readable explanations, making detection results more transparent and convincing. Our work sheds light on involving human factors to enhance defense against deepfakes.

📄 PDF Abstract BibTeX arXiv:2507.14807

Code (0)

등록된 구현이 없습니다.

Tasks

Face Detection

Similar Papers 제목 키워드 기반

Is Seeing Believing? Evaluating Human Sensitivity to Synthetic Video

2026-03-14 · David Wegmann, Emil Stevnsborg, Søren Knudsen, Luca Rossi 외 arxiv

Advances in machine learning have enabled the creation of realistic synthetic videos known as deepfakes. As deepfakes proliferate, concerns about rapid spread of disinformation and manipulation of public perception are m…

Seeing, Hearing, and Knowing Together: Multimodal Strategies in Deepfake Videos Detection

2026-02-01 · Chen Chen, Dion Hoe-Lian Goh arxiv

As deepfake videos become increasingly difficult for people to recognise, understanding the strategies humans use is key to designing effective media literacy interventions. We conducted a study with 195 participants bet…

SeeingSounds: Learning Audio-to-Visual Alignment via Text

2025-10-10 · Simone Carnemolla, Matteo Pennisi, Chiara Russo, Simone Palazzo 외 arxiv

We introduce SeeingSounds, a lightweight and modular framework for audio-to-image generation that leverages the interplay between audio, language, and vision-without requiring any paired audio-visual data or training on …

Image Generation

Beyond Seeing Is Believing: On Crowdsourced Detection of Audiovisual Deepfakes

2026-05-06 · Michael Soprano, Andrea Cioci, Stefano Mizzaro arxiv

Deepfakes are increasingly realistic and easy to produce, raising concerns about the reliability of human judgments in misinformation settings. We study audiovisual deepfake detection by measuring how consistently crowd …

DeepFake Detection

Detecting Deepfakes Without Seeing Any

2023-11-02 · Tal Reiss, Bar Cavia, Yedid Hoshen

Deepfake attacks, malicious manipulation of media containing people, are a serious concern for society. Conventional deepfake detection methods train supervised classifiers to distinguish real media from previously encou…

DeepFake DetectionFace SwappingFact CheckingFake News Detection