paper-with-me

홈 › Papers

Scene Text Image Super-resolution based on Text-conditional Diffusion Models

2023-11-16 · Chihiro Noguchi, Shun Fukuda, Masao Yamanaka

Scene Text Image Super-resolution (STISR) has recently achieved great success as a preprocessing method for scene text recognition. STISR aims to transform blurred and noisy low-resolution (LR) text images in real-world settings into clear high-resolution (HR) text images suitable for scene text recognition. In this study, we leverage text-conditional diffusion models (DMs), known for their impressive text-to-image synthesis capabilities, for STISR tasks. Our experimental results revealed that text-conditional DMs notably surpass existing STISR methods. Especially when texts from LR text images are given as input, the text-conditional DMs are able to produce superior quality super-resolution text images. Utilizing this capability, we propose a novel framework for synthesizing LR-HR paired text image datasets. This framework consists of three specialized text-conditional DMs, each dedicated to text image synthesis, super-resolution, and image degradation. These three modules are vital for synthesizing distinct LR and HR paired images, which are more suitable for training STISR methods. Our experiments confirmed that these synthesized image pairs significantly enhance the performance of STISR methods in the TextZoom evaluation.

📄 PDF Abstract BibTeX arXiv:2311.09759

Code (1)

toyotainfotech/stisr-tcdm 공식 구현 pytorch

Tasks

Image GenerationImage Super-ResolutionScene Text RecognitionSuper-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Scene Text Telescope: Text-Focused Scene Image Super-Resolution

2021-06-19 · CVPR 2021 1 · Jingye Chen, Bin Li, xiangyang xue

Image super-resolution, which is often regarded as a preprocessing procedure of scene text recognition, aims to recover the realistic features from a low-resolution text image. It has always been challenging due to l…

Image Super-ResolutionOptical Character Recognition (OCR)PositionScene Text Recognition+1

Text Prior Guided Scene Text Image Super-resolution

2021-06-29 · jianqi ma, Shi Guo, Lei Zhang

Scene text image super-resolution (STISR) aims to improve the resolution and visual quality of low-resolution (LR) scene text images, and consequently boost the performance of text recognition. However, most of existing …

Image Super-ResolutionSuper-Resolution

Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance

2025-10-24 · Minxing Luo, Linlong Fan, Wang Qiushi, Ge Wu 외 arxiv

Current image super-resolution methods show strong performance on natural images but distort text, creating a fundamental trade-off between image quality and textual readability. To address this, we introduce TIGER (Text…

Image Super-ResolutionImage Enhancement

Recognition-Guided Diffusion Model for Scene Text Image Super-Resolution

2023-11-22 · Yuxuan Zhou, Liangcai Gao, Zhi Tang, Baole Wei

Scene Text Image Super-Resolution (STISR) aims to enhance the resolution and legibility of text within low-resolution (LR) images, consequently elevating recognition accuracy in Scene Text Recognition (STR). Previous met…

DenoisingDiversityImage Super-ResolutionScene Text Recognition+1

Scene Text Image Super-Resolution in the Wild

2020-05-07 · ECCV 2020 8 · Wenjia Wang, Enze Xie, Xuebo Liu, Wenhai Wang 외

Low-resolution text images are often seen in natural scenes such as documents captured by mobile phones. Recognizing low-resolution text images is challenging because they lose detailed content information, leading to po…

Image Super-ResolutionSuper-Resolution