paper-with-me

홈 › Papers

Revisiting Tampered Scene Text Detection in the Era of Generative AI

2024-07-31 · Chenfan Qu, Yiwu Zhong, Fengjun Guo, Lianwen Jin

The rapid advancements of generative AI have fueled the potential of generative text image editing, meanwhile escalating the threat of misinformation spreading. However, existing forensics methods struggle to detect unseen forgery types that they have not been trained on, underscoring the need for a model capable of generalized detection of tampered scene text. To tackle this, we propose a novel task: open-set tampered scene text detection, which evaluates forensics models on their ability to identify both seen and previously unseen forgery types. We have curated a comprehensive, high-quality dataset, featuring the texts tampered by eight text editing models, to thoroughly assess the open-set generalization capabilities. Further, we introduce a novel and effective training paradigm that subtly alters the texture of selected texts within an image and trains the model to identify these regions. This approach not only mitigates the scarcity of high-quality training data but also enhances models' fine-grained perception and open-set generalization abilities. Additionally, we present DAF, a novel framework that improves open-set generalization by distinguishing between the features of authentic and tampered text, rather than focusing solely on the tampered text's features. Our extensive experiments validate the remarkable efficacy of our methods. For example, our zero-shot performance can even beat the previous state-of-the-art full-shot model by a large margin. Our dataset and code are available at https://github.com/qcf-568/OSTF.

📄 PDF Abstract BibTeX arXiv:2407.21422

Code (1)

qcf-568/ostf 공식 구현 pytorch

Tasks

MisinformationScene Text DetectionText Detection

Similar Papers 제목 키워드 기반

TextSleuth: A New Dataset and Baseline for Scene Text Manipulation Detection

2024-08-07 · Conference on Multimedia Information Processing and Retrieval 2024 8 · Abhineet Kumar Pandey, Ming-Ching Chang Xin, Li

With the rise of digital content on social media and the advancement of image editing tools, tampering with scene text has become a serious concern. Scene text manipulation detection (STMD) is a kind of image manipulatio…

Image ManipulationImage Manipulation Detection

TextSleuth: Towards Explainable Tampered Text Detection

2024-12-19 · Chenfan Qu, Jian Liu, Haoxing Chen, Baihan Yu 외

Recently, tampered text detection has attracted increasing attention due to its essential role in information security. Although existing methods can detect the tampered text region, the interpretation of such detection …

Domain GeneralizationOptical Character Recognition (OCR)Text Detection

TextShield-R1: Reinforced Reasoning for Tampered Text Detection

2026-02-23 · Chenfan Qu, Yiwu Zhong, Jian Liu, Xuekang Zhu 외 arxiv

The growing prevalence of tampered images poses serious security threats, highlighting the urgent need for reliable detection methods. Multimodal large language models (MLLMs) demonstrate strong potential in analyzing ta…

Reinforcement LearningText Detection

Towards Robust Tampered Text Detection in Document Image: New Dataset and New Solution

2023-01-01 · CVPR 2023 1 · Chenfan Qu, Chongyu Liu, Yuliang Liu, Xinhong Chen 외

Recently, tampered text detection in document image has attracted increasingly attention due to its essential role on information security. However, detecting visually consistent tampered text in photographed documen…

DecoderImage and Video Forgery DetectionImage CompressionText Detection

SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal Model

2024-12-05 · CVPR 2025 1 · Zhenglin Huang, Jinwei Hu, Xiangtai Li, Yiwei He 외

The rapid advancement of generative models in creating highly realistic images poses substantial risks for misinformation dissemination. For instance, a synthetic image, when shared on social media, can mislead extensive…

DeepFake DetectionFace SwappingMisinformation