paper-with-me

홈 › Papers

ForensicZip: More Tokens are Better but Not Necessary in Forensic Vision-Language Models

2026-03-12 · Yingxin Lai, Zitong Yu, Jun Wang, Linlin Shen, Yong Xu, Xiaochun Cao arxiv

Multimodal Large Language Models (MLLMs) enable interpretable multimedia forensics by generating textual rationales for forgery detection. However, processing dense visual sequences incurs high computational costs, particularly for high-resolution images and videos. Visual token pruning is a practical acceleration strategy, yet existing methods are largely semantic-driven, retaining salient objects while discarding background regions where manipulation traces such as high-frequency anomalies and temporal jitters often reside. To address this issue, we introduce ForensicZip, a training-free framework that reformulates token compression from a forgery-driven perspective. ForensicZip models temporal token evolution as a Birth-Death Optimal Transport problem with a slack dummy node, quantifying physical discontinuities indicating transient generative artifacts. The forensic scoring further integrates transport-based novelty with high-frequency priors to separate forensic evidence from semantic content under large-ratio compression. Experiments on deepfake and AIGC benchmarks show that at 10\% token retention, ForensicZip achieves $2.97\times$ speedup and over 90\% FLOPs reduction while maintaining state-of-the-art detection performance.

📄 PDF Abstract BibTeX arXiv:2603.12208

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ForensicsData: A Digital Forensics Dataset for Large Language Models

2025-08-31 · Youssef Chakir, Iyad Lahsen-Cherif arxiv

The growing complexity of cyber incidents presents significant challenges for digital forensic investigators, especially in evidence collection and analysis. Public resources are still limited because of ethical, legal, …

ForensicsTok: Forensics-Guided Tokenized Modeling for Image Tampering Localization

2026-06-23 · Lei Xu, Haowei Wang, Shen Chen, Taiping Yao 외 arxiv

Multi-modal Large Language Models (MLLMs) offer powerful reasoning for forensic tasks, yet existing approaches utilizing exogenous segmentation decoders often suffer from suboptimal localization. The reliance on stitched…

Image Manipulation Localization

Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection

2026-05-25 · Izaldein Al-Zyoud, Abdulmotaleb El Saddik arxiv

We introduce segmentation-guided spatial indexing for generalizable and explainable deepfake detection. The key idea reverses the standard design order: rather than pooling all facial tokens and classifying afterward, we…

DeepFake Detection

Digital Image Forensics: A quantitative & qualitative comparison between State-of-the-art-AI and Traditional Techniques for detection and localization of image manipulations

2024-09-01 · None 2024 9 · UHstudent

With the rise of realistic AI-generated images and continuously advancing photo-editing software, it has become increasingly difficult to reliably distinguish between authentic and manipulated images. Using Digital Image…

Detecting Image ManipulationImage ForensicsImage Forgery DetectionImage Manipulation+4

The Forensic Cost of Watermark Removal: From Dedicated Attacks to Image Editing

2026-04-28 · Gautier Evennou, Ewa Kijak arxiv

Current watermark removal methods are evaluated on two axes: attack success rate and perceptual quality. We show this is insufficient. While state-of-the-art attacks successfully degrade the watermark signal without visi…

Image Editing