paper-with-me

Papers

SocialDF: Benchmark Dataset and Detection Model for Mitigating Harmful Deepfake Content on Social Media Platforms

2025-06-05 · Arnesh Batra, Anushk Kumar, Jashn Khemani, Arush Gumber, Arhan Jain, Somil Gupta

The rapid advancement of deep generative models has significantly improved the realism of synthetic media, presenting both opportunities and security challenges. While deepfake technology has valuable applications in entertainment and accessibility, it has emerged as a potent vector for misinformation campaigns, particularly on social media. Existing detection frameworks struggle to distinguish between benign and adversarially generated deepfakes engineered to manipulate public perception. To address this challenge, we introduce SocialDF, a curated dataset reflecting real-world deepfake challenges on social media platforms. This dataset encompasses high-fidelity deepfakes sourced from various online ecosystems, ensuring broad coverage of manipulative techniques. We propose a novel LLM-based multi-factor detection approach that combines facial recognition, automated speech transcription, and a multi-agent LLM pipeline to cross-verify audio-visual cues. Our methodology emphasizes robust, multi-modal verification techniques that incorporate linguistic, behavioral, and contextual analysis to effectively discern synthetic media from authentic content.

📄 PDF Abstract BibTeX arXiv:2506.05538

Code (1)

arnesh2212/SocialDF 공식 구현

Tasks

Face SwappingMisinformation

Similar Papers 제목 키워드 기반

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval

2025-10-10 · Wenshuo Wang, Ziyou Jiang, Junjie Wang, Mingyang Li 외 arxiv

Internet memes have emerged as a popular multimodal medium, yet they are increasingly weaponized to convey harmful opinions through subtle rhetorical devices like irony and metaphor. Existing detection approaches, includ…

Deepfake Detection: A Comparative Analysis

2023-08-07 · Sohail Ahmed Khan, Duc-Tien Dang-Nguyen

This paper present a comprehensive comparative analysis of supervised and self-supervised models for deepfake detection. We evaluate eight supervised deep learning architectures and two transformer-based models pre-train…

DeepFake DetectionDeep LearningFace Swapping

HOD: A Benchmark Dataset for Harmful Object Detection

2023-10-08 · Eungyeom Ha, Heemook Kim, Sung Chul Hong, Dongbin Na

Recent multi-media data such as images and videos have been rapidly spread out on various online services such as social network services (SNS). With the explosive growth of online media services, the number of image con…

Objectobject-detectionObject Detection

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

2026-04-18 · Huije Lee, Jisu Shin, Hoyun Song, Changgeon Ko 외 arxiv

Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale pre-training corpora. To address these issues, we propose a framewor…

Mitigating harm in language models with conditional-likelihood filtration

2021-08-04 · Helen Ngo, Cooper Raterink, João G. M. Araújo, Ivan Zhang 외

Language models trained on large-scale unfiltered datasets curated from the open web acquire systemic biases, prejudices, and harmful views from their training data. We present a methodology for programmatically identify…

Language ModelingLanguage Modelling