paper-with-me

홈 › Papers

Grab-3D: Detecting AI-Generated Videos from 3D Geometric Temporal Consistency

2025-12-15 · Wenhan Chen, Sezer Karaoglu, Theo Gevers arxiv

Recent advances in diffusion-based generation techniques enable AI models to produce highly realistic videos, heightening the need for reliable detection mechanisms. However, existing detection methods provide only limited exploration of the 3D geometric patterns present in generated videos. In this paper, we use vanishing points as an explicit representation of 3D geometry patterns, revealing fundamental discrepancies in geometric consistency between real and AI-generated videos. We introduce Grab-3D, a geometry-aware transformer framework for detecting AI-generated videos based on 3D geometric temporal consistency. To enable reliable evaluation, we construct an AI-generated video dataset of static scenes, allowing stable 3D geometric feature extraction. We propose a geometry-aware transformer equipped with geometric positional encoding, temporal-geometric attention, and an EMA-based geometric classifier head to explicitly inject 3D geometric awareness into temporal modeling. Experiments demonstrate that Grab-3D significantly outperforms state-of-the-art detectors, achieving robust cross-domain generalization to unseen generators.

📄 PDF Abstract BibTeX arXiv:2512.13665

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Generalization

Similar Papers 제목 키워드 기반

Improving the Efficiency and Robustness of Deepfakes Detection through Precise Geometric Features

2021-04-09 · CVPR 2021 1 · Zekun Sun, Yujie Han, Zeyu Hua, Na Ruan 외

Deepfakes is a branch of malicious techniques that transplant a target face to the original one in videos, resulting in serious problems such as infringement of copyright, confusion of information, or even public panic. …

Open-Ended Question Answering

MotionPhys: Detecting AI-Generated Videos via Physical Consistency of Optical-Flow Trajectories

2026-08-21 · Haojin He, Hao Tan, Zichang Tan, Ajian Liu 외 arxiv

Modern AI video generation models can produce videos with high visual fidelity and seemingly smooth temporal transitions. However, visual realism does not necessarily imply physical motion consistency. Existing generativ…

Video Generation

Turns Out I'm Not Real: Towards Robust Detection of AI-Generated Videos

2024-06-13 · Qingyuan Liu, Pengyuan Shi, Yun-Yun Tsai, Chengzhi Mao 외

The impressive achievements of generative models in creating high-quality videos have raised concerns about digital integrity and privacy vulnerabilities. Recent works to combat Deepfakes videos have developed detectors …

A Proposal-Based Solution to Spatio-Temporal Action Detection in Untrimmed Videos

2018-11-20 · Joshua Gleason, Rajeev Ranjan, Steven Schwarcz, Carlos D. Castillo 외

Existing approaches for spatio-temporal action detection in videos are limited by the spatial extent and temporal duration of the actions. In this paper, we present a modular system for spatio-temporal action detection i…

Action ClassificationAction DetectionClusteringobject-detection+1

Consolidating Diffusion-Generated Video Detection with Unified Multimodal Forgery Learning

2025-11-22 · Xiaohong Liu, Xiufeng Song, Huayu Zheng, Lei Bai 외 arxiv

The proliferation of videos generated by diffusion models has raised increasing concerns about information security, highlighting the urgent need for reliable detection of synthetic media. Existing methods primarily focu…