paper-with-me

홈 › Papers

SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection

2026-05-17 · Zixi Wei, Huixuaun Zhang, Xiaojun Wan arxiv

The remarkable visual fidelity of recent commercial video generative models, such as Sora and Veo, renders robust AI-generated video detection increasingly essential to prevent synthetic content from being indistinguishable from real videos and exploited for disinformation. However, existing detectors often fail due to an over-reliance on increasingly realistic semantic features, neglecting subtle spectral artifacts. In this paper, we propose SpecSem-Net, the first framework to introduce a semantic-guided spectral denoising mechanism specifically for high-fidelity AI-generated video detection. Specifically, we design a spectral module to extract high-frequency features via Fourier-Transform based filtering. Furthermore, to reduce misjudgments arising from spectral noise, we employ a Gated Merging Mechanism to adaptively fuse semantic context, effectively mitigating spectral noise. Additionally, to evaluate detector performance on the latest top-tier generative models, we construct a comprehensive benchmark comprising 5 SOTA commercial generators. Extensive experiments demonstrate that SpecSem-Net outperforms existing methods, achieving accuracies of 87.25% and 95.59% on our benchmark and public datasets, respectively.

📄 PDF Abstract BibTeX arXiv:2605.17311

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FUSE: Unifying Spectral and Semantic Cues for Robust AI-Generated Image Detection

2025-12-25 · Md. Zahid Hossain, Most. Sharmin Sultana Samu, Md. Kamrozzaman Bhuiyan, Farhad Uz Zaman 외 arxiv

The fast evolution of generative models has heightened the demand for reliable detection of AI-generated images. To tackle this challenge, we introduce FUSE, a hybrid system that combines spectral features extracted thro…

Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection

2024-04-26 · Jiawei Song, Dengpan Ye, Yunming Zhang

Artificial Intelligence Generated Content (AIGC) techniques, represented by text-to-image generation, have led to a malicious use of deep forgeries, raising concerns about the trustworthiness of multimedia content. Adapt…

Image GenerationText to Image GenerationText-to-Image Generation

Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation

2025-05-21 · Ce Zhang, Zifu Wan, Simon Stepputtis, Katia Sycara 외

Semantic segmentation relying solely on RGB data often struggles in challenging conditions such as low illumination and obscured views, limiting its reliability in critical applications like autonomous driving. To addres…

Autonomous DrivingSemantic Segmentation

Reliable Explainability of Deep Learning Spatial-Spectral Classifiers for Improved Semantic Segmentation in Autonomous Driving

2025-02-20 · Jon Gutiérrez-Zaballa, Koldo Basterretxea, Javier Echanobe

Integrating hyperspectral imagery (HSI) with deep neural networks (DNNs) can strengthen the accuracy of intelligent vision systems by combining spectral and spatial information, which is useful for tasks like semantic se…

Autonomous Drivingimage-classificationImage ClassificationSemantic Segmentation

Exploring Multi-Timestep Multi-Stage Diffusion Features for Hyperspectral Image Classification

2023-06-15 · Jingyi Zhou, Jiamu Sheng, Jiayuan Fan, Peng Ye 외

The effectiveness of spectral-spatial feature learning is crucial for the hyperspectral image (HSI) classification task. Diffusion models, as a new class of groundbreaking generative models, have the ability to learn bot…

ClassificationHyperspectral Image Classificationimage-classificationImage Classification