paper-with-me

홈 › Papers

Detecting Multimedia Generated by Large AI Models: A Survey

2024-01-22 · Li Lin, Neeraj Gupta, Yue Zhang, Hainan Ren, Chun-Hao Liu, Feng Ding, Xin Wang, Xin Li, Luisa Verdoliva, Shu Hu

The rapid advancement of Large AI Models (LAIMs), particularly diffusion models and large language models, has marked a new era where AI-generated multimedia is increasingly integrated into various aspects of daily life. Although beneficial in numerous fields, this content presents significant risks, including potential misuse, societal disruptions, and ethical concerns. Consequently, detecting multimedia generated by LAIMs has become crucial, with a marked rise in related research. Despite this, there remains a notable gap in systematic surveys that focus specifically on detecting LAIM-generated multimedia. Addressing this, we provide the first survey to comprehensively cover existing research on detecting multimedia (such as text, images, videos, audio, and multimodal content) created by LAIMs. Specifically, we introduce a novel taxonomy for detection methods, categorized by media modality, and aligned with two perspectives: pure detection (aiming to enhance detection performance) and beyond detection (adding attributes like generalizability, robustness, and interpretability to detectors). Additionally, we have presented a brief overview of generation mechanisms, public datasets, online detection tools, and evaluation metrics to provide a valuable resource for researchers and practitioners in this field. Most importantly, we offer a focused analysis from a social media perspective to highlight their broader societal impact. Furthermore, we identify current challenges in detection and propose directions for future research that address unexplored, ongoing, and emerging issues in detecting multimedia generated by LAIMs. Our aim for this survey is to fill an academic gap and contribute to global AI security efforts, helping to ensure the integrity of information in the digital realm. The project link is https://github.com/Purdue-M2/Detect-LAIM-generated-Multimedia-Survey.

📄 PDF Abstract BibTeX arXiv:2402.00045

Code (1)

purdue-m2/detect-laim-generated-multimedia-survey 공식 구현

Tasks

Survey

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Focus 설명 없음

Similar Papers 제목 키워드 기반

Affective Computing for Large-Scale Heterogeneous Multimedia Data: A Survey

2019-10-03 · Sicheng Zhao, Shangfei Wang, Mohammad Soleymani, Dhiraj Joshi 외

The wide popularity of digital photography and social networks has generated a rapidly growing volume of multimedia data (i.e., image, music, and video), resulting in a great demand for managing, retrieving, and understa…

Survey

Methods and Trends in Detecting Generated Images: A Comprehensive Review

2025-02-21 · Arpan Mahara, Naphtali Rishe

The proliferation of generative models, such as Generative Adversarial Networks (GANs), Diffusion Models, and Variational Autoencoders (VAEs), has enabled the synthesis of high-quality multimedia data. However, these adv…

BenchmarkingDeepFake DetectionFace SwappingSynthetic Image Detection

A Survey of Multimedia Technologies and Robust Algorithms

2021-03-24 · Zijian Kuang, Xinran Tie

Multimedia technologies are now more practical and deployable in real life, and the algorithms are widely used in various researching areas such as deep learning, signal processing, haptics, computer vision, robotics, an…

Survey

SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs

2025-11-16 · Shail Desai, Aditya Pawar, Li Lin, Xin Wang 외 arxiv

Artificial Intelligence (AI) has made it possible for anyone to create images, audio, and video with unprecedented ease, enriching education, communication, and creative expression. At the same time, the rapid rise of AI…

DeepFake Detection

Datasets, Clues and State-of-the-Arts for Multimedia Forensics: An Extensive Review

2024-01-13 · Ankit Yadav, Dinesh Kumar Vishwakarma

With the large chunks of social media data being created daily and the parallel rise of realistic multimedia tampering methods, detecting and localising tampering in images and videos has become essential. This survey fo…

DeepFake DetectionDeep LearningFace Swapping