paper-with-me

Papers

SDiFL: Stable Diffusion-Driven Framework for Image Forgery Localization

2025-08-27 · Yang Su, Shunquan Tan, Jiwu Huang arxiv

Driven by the new generation of multi-modal large models, such as Stable Diffusion (SD), image manipulation technologies have advanced rapidly, posing significant challenges to image forensics. However, existing image forgery localization methods, which heavily rely on labor-intensive and costly annotated data, are struggling to keep pace with these emerging image manipulation technologies. To address these challenges, we are the first to integrate both image generation and powerful perceptual capabilities of SD into an image forensic framework, enabling more efficient and accurate forgery localization. First, we theoretically show that the multi-modal architecture of SD can be conditioned on forgery-related information, enabling the model to inherently output forgery localization results. Then, building on this foundation, we specifically leverage the multimodal framework of Stable DiffusionV3 (SD3) to enhance forgery localization performance.We leverage the multi-modal processing capabilities of SD3 in the latent space by treating image forgery residuals -- high-frequency signals extracted using specific highpass filters -- as an explicit modality. This modality is fused into the latent space during training to enhance forgery localization performance. Notably, our method fully preserves the latent features extracted by SD3, thereby retaining the rich semantic information of the input image. Experimental results show that our framework achieves up to 12% improvements in performance on widely used benchmarking datasets compared to current state-of-the-art image forgery localization models. Encouragingly, the model demonstrates strong performance on forensic tasks involving real-world document forgery images and natural scene forging images, even when such data were entirely unseen during training.

📄 PDF Abstract BibTeX arXiv:2508.20182

Code (0)

등록된 구현이 없습니다.

Tasks

Image ManipulationImage Generation

Similar Papers 제목 키워드 기반

StableVideo: Text-driven Consistency-aware Diffusion Video Editing

2023-08-18 · ICCV 2023 1 · Wenhao Chai, Xun Guo, Gaoang Wang, Yan Lu

Diffusion-based methods can generate realistic images and videos, but they struggle to edit existing objects in a video while preserving their appearance over time. This prevents diffusion models from being applied to na…

Video Editing

StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation

2025-07-20 · Shuyuan Tu, Zhen Xing, Xintong Han, Zhi-Qi Cheng 외 arxiv

Current diffusion models for human image animation often struggle to maintain identity (ID) consistency, especially when the reference image and driving video differ significantly in body size or position. We introduce S…

GUNet: A Graph Convolutional Network United Diffusion Model for Stable and Diversity Pose Generation

2024-09-18 · Shuowen Liang, Sisi Li, Qingyun Wang, Cen Zhang 외

Pose skeleton images are an important reference in pose-controllable image generation. In order to enrich the source of skeleton images, recent works have investigated the generation of pose skeletons based on natural la…

DenoisingDiversityImage Generation

TF-ICON: Diffusion-Based Training-Free Cross-Domain Image Composition

2023-07-24 · ICCV 2023 1 · Shilin Lu, Yanzhu Liu, Adams Wai-Kin Kong

Text-driven diffusion models have exhibited impressive generative capabilities, enabling various image editing tasks. In this paper, we propose TF-ICON, a novel Training-Free Image COmpositioN framework that harnesses th…

Image-Guided CompositionText-to-Image Generation

Uni-paint: A Unified Framework for Multimodal Image Inpainting with Pretrained Diffusion Model

2023-10-11 · Shiyuan Yang, Xiaodong Chen, Jing Liao

Recently, text-to-image denoising diffusion probabilistic models (DDPMs) have demonstrated impressive image generation capabilities and have also been successfully applied to image inpainting. However, in practice, users…

DenoisingImage DenoisingImage GenerationImage Inpainting