paper-with-me

홈 › Papers

DiffSeg30k: A Multi-Turn Diffusion Editing Benchmark for Localized AIGC Detection

2025-11-24 · Hai Ci, Ziheng Peng, Pei Yang, Yingxin Xuan, Mike Zheng Shou arxiv

Diffusion-based editing enables realistic modification of local image regions, making AI-generated content harder to detect. Existing AIGC detection benchmarks focus on classifying entire images, overlooking the localization of diffusion-based edits. We introduce DiffSeg30k, a publicly available dataset of 30k diffusion-edited images with pixel-level annotations, designed to support fine-grained detection. DiffSeg30k features: 1) In-the-wild images--we collect images or image prompts from COCO to reflect real-world content diversity; 2) Diverse diffusion models--local edits using eight SOTA diffusion models; 3) Multi-turn editing--each image undergoes up to three sequential edits to mimic real-world sequential editing; and 4) Realistic editing scenarios--a vision-language model (VLM)-based pipeline automatically identifies meaningful regions and generates context-aware prompts covering additions, removals, and attribute changes. DiffSeg30k shifts AIGC detection from binary classification to semantic segmentation, enabling simultaneous localization of edits and identification of the editing models. We benchmark three baseline segmentation approaches, revealing significant challenges in semantic segmentation tasks, particularly concerning robustness to image distortions. Experiments also reveal that segmentation models, despite being trained for pixel-level localization, emerge as highly reliable whole-image classifiers of diffusion edits, outperforming established forgery classifiers while showing great potential in cross-generator generalization. We believe DiffSeg30k will advance research in fine-grained localization of AI-generated content by demonstrating the promise and limitations of segmentation-based methods. DiffSeg30k is released at: https://huggingface.co/datasets/Chaos2629/Diffseg30k

📄 PDF Abstract BibTeX arXiv:2511.19111

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic SegmentationBinary Classification

Similar Papers 제목 키워드 기반

TextDiffSeg: Text-guided Latent Diffusion Model for 3d Medical Images Segmentation

2025-04-16 · Kangbo Ma

Diffusion Probabilistic Models (DPMs) have demonstrated significant potential in 3D medical image segmentation tasks. However, their high computational cost and inability to fully capture global 3D contextual information…

Image SegmentationLatent Diffusion Model for 3DMedical Image SegmentationOrgan Segmentation+2

PGDiffSeg: Prior-Guided Denoising Diffusion Model with Parameter-Shared Attention for Breast Cancer Segmentation

2024-10-23 · Feiyan Feng, Tianyu Liu, Hong Wang, Jun Zhao 외

Early detection through imaging and accurate diagnosis is crucial in mitigating the high mortality rate associated with breast cancer. However, locating tumors from low-resolution and high-noise medical images is extreme…

DenoisingImage SegmentationMedical Image SegmentationSegmentation+1

DiffSeg: A Segmentation Model for Skin Lesions Based on Diffusion Difference

2024-04-25 · Zhihao Shuai, Yinan Chen, Shunqiang Mao, Yihan Zho 외

Weakly supervised medical image segmentation (MIS) using generative models is crucial for clinical diagnosis. However, the accuracy of the segmentation results is often limited by insufficient supervision and the complex…

Decision MakingImage SegmentationMedical Image SegmentationSegmentation+1

DiffSegLung: Diffusion Radiomic Distillation for Unsupervised Lung Pathology Segmentation

2026-05-12 · Rezkellah Noureddine Khiati, Pierre-Yves Brillet, Catalin Fetita arxiv

Unsupervised segmentation of pulmonary pathologies in CT remains an open challenge due to the absence of annotated multi pathology cohorts and the failure of existing diffusion-based methods to exploit the quantitative H…

VINCIE: Unlocking In-context Image Editing from Video

2025-06-12 · Leigang Qu, Feng Cheng, Ziyan Yang, Qi Zhao 외

In-context image editing aims to modify images based on a contextual sequence comprising text and previously generated images. Existing methods typically depend on task-specific pipelines and expert models (e.g., segment…

PredictionSegmentationStory Generation