paper-with-me

홈 › Papers

FUSED: Forensic-Semantic Mixture-of-Experts for AI Inpainting Detection and Localization

2026-08-28 · Anton Nuzhdin, Marcel Worring, Ivona Najdenkoska arxiv

Diffusion-based inpainting models modify only a localized part of an image, while many AI-image detectors rely on global artifacts and do not localize. These artifacts vary across generators, limiting detector transfer under distribution shifts. Recent work shows that restoring the authentic pixels outside the inpainted region removes these cues and can degrade pretrained detectors. To address this, we present FUSED, a unified framework for the joint detection and localization of AI-generated inpainting. FUSED combines low-level forensic cues with high-level semantic features using a sparsely-gated Mixture-of-Experts architecture, enabling the model to adaptively prioritize the most relevant signal for each token. For each input, FUSED predicts both an image-level manipulation score and a pixel-level mask of the inpainted area. On the OpenSDID cross-generator benchmark, FUSED achieves the best average detection and localization, with the largest gains on unseen generators. The same model transfers directly to the held-out AutoSplice and CocoGlide benchmarks, more than doubling localization performance. Evaluating each held-out benchmark with and without the global generator artifact further shows that all evaluated methods, ours included, partly read the artifact as evidence of manipulation, and FUSED remains the strongest under both conditions. Code and pretrained models are available at https://github.com/AntonNuzhdin/FUSED.

📄 PDF Abstract BibTeX arXiv:2608.28302

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Semantic-Aware Dynamic Parameter for Video Inpainting Transformer

2023-01-01 · ICCV 2023 1 · Eunhye Lee, Jinsu Yoo, Yunjeong Yang, Sungyong Baik 외

Recent learning-based video inpainting approaches have achieved considerable progress. However, they still cannot fully utilize semantic information within the video frames and predict improper scene layout, failing …

Mixture-of-ExpertsVideo Inpainting

SafePaint: Anti-forensic Image Inpainting with Domain Adaptation

2024-04-28 · Dunyun Chen, Xin Liao, Xiaoshuai Wu, Shiwei Chen

Existing image inpainting methods have achieved remarkable accomplishments in generating visually appealing results, often accompanied by a trend toward creating more intricate structural textures. However, while these m…

Domain AdaptationImage Inpainting

TGIF2: Extended Text-Guided Inpainting Forgery Dataset & Benchmark

2026-03-30 · Hannes Mareen, Dimitrios Karageorgiou, Paschalis Giakoumoglou, Peter Lambert 외 arxiv

Generative AI has made text-guided inpainting a powerful image editing tool, but at the same time a growing challenge for media forensics. Existing benchmarks, including our text-guided inpainting forgery (TGIF) dataset,…

Image EnhancementImage Editing

UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts

2024-12-09 · Zhen Wan, Yue Ma, Chenyang Qi, Zhiheng Liu 외

In this paper, we present UniPaint, a unified generative space-time video inpainting framework that enables spatial-temporal inpainting and interpolation. Different from existing methods that treat video inpainting and v…

Mixture-of-ExpertsVideo Inpainting

StableGuard: Towards Unified Copyright Protection and Tamper Localization in Latent Diffusion Models

2025-09-22 · Haoxin Yang, Bangzhen Liu, Xuemiao Xu, Cheng Xu 외 arxiv

The advancement of diffusion models has enhanced the realism of AI-generated content but also raised concerns about misuse, necessitating robust copyright protection and tampering localization. Although recent methods ha…