paper-with-me

홈 › Papers

GenWorld: Towards Detecting AI-generated Real-world Simulation Videos

2025-06-12 · Weiliang Chen, Wenzhao Zheng, Yu Zheng, Lei Chen, Jie zhou, Jiwen Lu, Yueqi Duan

The flourishing of video generation technologies has endangered the credibility of real-world information and intensified the demand for AI-generated video detectors. Despite some progress, the lack of high-quality real-world datasets hinders the development of trustworthy detectors. In this paper, we propose GenWorld, a large-scale, high-quality, and real-world simulation dataset for AI-generated video detection. GenWorld features the following characteristics: (1) Real-world Simulation: GenWorld focuses on videos that replicate real-world scenarios, which have a significant impact due to their realism and potential influence; (2) High Quality: GenWorld employs multiple state-of-the-art video generation models to provide realistic and high-quality forged videos; (3) Cross-prompt Diversity: GenWorld includes videos generated from diverse generators and various prompt modalities (e.g., text, image, video), offering the potential to learn more generalizable forensic features. We analyze existing methods and find they fail to detect high-quality videos generated by world models (i.e., Cosmos), revealing potential drawbacks of ignoring real-world clues. To address this, we propose a simple yet effective model, SpannDetector, to leverage multi-view consistency as a strong criterion for real-world AI-generated video detection. Experiments show that our method achieves superior results, highlighting a promising direction for explainable AI-generated video detection based on physical plausibility. We believe that GenWorld will advance the field of AI-generated video detection. Project Page: https://chen-wl20.github.io/GenWorld

📄 PDF Abstract BibTeX arXiv:2506.10975

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks

2026-03-29 · Samin Mahdizadeh Sani, Max Ku, Nima Jamali, Matina Mahdizadeh Sani 외 arxiv

Advances in diffusion, autoregressive, and hybrid models have enabled high-quality image synthesis for tasks such as text-to-image, editing, and reference-guided composition. Yet, existing benchmarks remain limited, eith…

Image Generation

Learning from synthetic data generated with GRADE

2023-05-07 · Elia Bonetto, Chenghao Xu, Aamir Ahmad

Recently, synthetic data generation and realistic rendering has advanced tasks like target tracking and human pose estimation. Simulations for most robotics applications are obtained in (semi)static environments, with sp…

Pose EstimationSynthetic Data Generation

Detecting CNN-Generated Facial Images in Real-World Scenarios

2020-05-12 · Nils Hulzebosch, Sarah Ibrahimi, Marcel Worring

Artificial, CNN-generated images are now of such high quality that humans have trouble distinguishing them from real images. Several algorithmic detection methods have been proposed, but these appear to generalize poorly…

Detecting Affordances by Visuomotor Simulation

2016-11-01 · Wolfram Schenck, Hendrik Hasenbein, Ralf Möller

The term "affordance" denotes the behavioral meaning of objects. We propose a cognitive architecture for the detection of affordances in the visual modality. This model is based on the internal simulation of movement seq…

Affordance Detection

Detecting Images Generated by Diffusers

2023-03-09 · Davide Alessandro Coccomini, Andrea Esuli, Fabrizio Falchi, Claudio Gennaro 외

This paper explores the task of detecting images generated by text-to-image diffusion models. To evaluate this, we consider images generated from captions in the MSCOCO and Wikimedia datasets using two state-of-the-art m…