paper-with-me

Papers

CoDA: Color Distribution Probing for Efficient and Generalizable AI-Generated Image Detection

2026-05-23 · Zexi Jia, Zhiqiang Yuan, Xiaoyue Duan, Jinchao Zhang, Jie Zhou, Anil K. Jain arxiv

AI-generated image detection faces a persistent trade-off between generalization and efficiency: lightweight artifact-based methods often degrade on unseen generators or domains, whereas more robust large-scale models are computationally expensive. Meanwhile, existing benchmarks mainly focus on cross-model evaluation in photorealistic settings, leaving cross-domain robustness underexplored. To address this gap, we introduce FakeForm, a large-scale benchmark with approximately 370,000 images across 62 diverse domains for both cross-model and cross-domain evaluation. Motivated by this broader setting, we revisit color-distribution probing as an efficient complementary cue for AI-generated image detection. We observe that, especially for photographic content, real photographs tend to exhibit smoother and more stable color patterns, whereas synthetic images often show characteristic color imbalances introduced by neural generation. Based on this observation, we propose CoDA, a compact 1.48M-parameter detector built on a Noise-Quantization Probe, together with a theoretical analysis linking probe responses to color non-uniformity. Experiments show that CoDA achieves state-of-the-art performance on standard benchmarks and the best results on the challenging cross-domain evaluation of FakeForm, while remaining highly competitive in cross-model photorealistic settings. These results suggest that persistent generative artifacts can provide a practical foundation for efficient and robust AI-generated image detection. The models and FakeForm benchmark will be made publicly available.

📄 PDF Abstract BibTeX arXiv:2605.24306

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The World of an Octopus: How Reporting Bias Influences a Language Model's Perception of Color

2021-10-15 · Cory Paik, Stéphane Aroca-Ouellette, Alessandro Roncone, Katharina Kann

Recent work has raised concerns about the inherent limitations of text-only pretraining. In this paper, we first demonstrate that reporting bias, the tendency of people to not state the obvious, is one of the causes of t…

Language ModelingLanguage Modelling

The World of an Octopus: How Reporting Bias Influences a Language Model’s Perception of Color

2021-11-01 · EMNLP 2021 11 · Cory Paik, Stéphane Aroca-Ouellette, Alessandro Roncone, Katharina Kann

Recent work has raised concerns about the inherent limitations of text-only pretraining. In this paper, we first demonstrate that reporting bias, the tendency of people to not state the obvious, is one of the causes of t…

Language ModelingLanguage Modelling

Probing in the Wild: A Case Study of Self-Supervised Speech Representations on Mandarin Sub-dialects with Unsupervised Articulatory Analysis

2026-06-24 · Shu Shang, Fuliang Weng, Zeqian Hu, Yaqian Zhou arxiv

While self-supervised speech models have achieved strong performance across speech tasks, relatively little is known about how their internal phonetic representations behave under fine-grained dialect variation. Existing…

Color Matters: Demosaicing-Guided Color Correlation Training for Generalizable AI-Generated Image Detection

2026-01-30 · Nan Zhong, Yiran Xu, Mian Zou arxiv

As realistic AI-generated images threaten digital authenticity, we address the generalization failure of generative artifact-based detectors by exploiting the intrinsic properties of the camera imaging pipeline. Concrete…

Cross-Modal-Domain Generalization Through Semantically Aligned Discrete Representations

2026-05-12 · Souptik Sen, Raneen Younis, Zahra Ahmadi arxiv

Multimodal learning seeks to integrate information across diverse sensory sources, yet current approaches struggle to balance cross-modal generalizability with modality-specific structure. Continuous (implicit) methods p…

Representation LearningDomain GeneralizationVideo Segmentation