paper-with-me

Papers

Boosting Diffusion-Based Text Image Super-Resolution Model Towards Generalized Real-World Scenarios

2025-03-10 · Chenglu Pan, Xiaogang Xu, Ganggui Ding, Yunke Zhang, Wenbo Li, Jiarong Xu, Qingbiao Wu

Restoring low-resolution text images presents a significant challenge, as it requires maintaining both the fidelity and stylistic realism of the text in restored images. Existing text image restoration methods often fall short in hard situations, as the traditional super-resolution models cannot guarantee clarity, while diffusion-based methods fail to maintain fidelity. In this paper, we introduce a novel framework aimed at improving the generalization ability of diffusion models for text image super-resolution (SR), especially promoting fidelity. First, we propose a progressive data sampling strategy that incorporates diverse image types at different stages of training, stabilizing the convergence and improving the generalization. For the network architecture, we leverage a pre-trained SR prior to provide robust spatial reasoning capabilities, enhancing the model's ability to preserve textual information. Additionally, we employ a cross-attention mechanism to better integrate textual priors. To further reduce errors in textual priors, we utilize confidence scores to dynamically adjust the importance of textual features during training. Extensive experiments on real-world datasets demonstrate that our approach not only produces text images with more realistic visual appearances but also improves the accuracy of text structure.

📄 PDF Abstract BibTeX arXiv:2503.07232

Code (0)

등록된 구현이 없습니다.

Tasks

Image RestorationImage Super-ResolutionSpatial ReasoningSuper-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

PEAN: A Diffusion-Based Prior-Enhanced Attention Network for Scene Text Image Super-Resolution

2023-11-29 · Zuoyan Zhao, Hui Xue, Pengfei Fang, Shipeng Zhu

Scene text image super-resolution (STISR) aims at simultaneously increasing the resolution and readability of low-resolution scene text images, thus boosting the performance of the downstream recognition task. Two factor…

Image Super-ResolutionMulti-Task LearningSuper-Resolution

Dreamix: Video Diffusion Models are General Video Editors

2023-02-02 · Eyal Molad, Eliahu Horwitz, Dani Valevski, Alex Rav Acha 외

Text-driven image and video diffusion models have recently achieved unprecedented generation realism. While diffusion models have been successfully applied for image editing, very few works have done so for video editing…

Image AnimationImage to Video GenerationSubject-driven Video GenerationText-to-Video Editing+2

RealOSR: Latent Unfolding Boosting Diffusion-based Real-world Omnidirectional Image Super-Resolution

2024-12-11 · Xuhan Sheng, Runyi Li, Bin Chen, Weiqi Li 외

Omnidirectional image super-resolution (ODISR) aims to upscale low-resolution (LR) omnidirectional images (ODIs) to high-resolution (HR), addressing the growing demand for detailed visual content across a $180^{\circ}\ti…

DenoisingImage Super-ResolutionSuper-Resolution

Boosting Diffusion Guidance via Learning Degradation-Aware Models for Blind Super Resolution

2025-01-15 · Shao-Hao Lu, Ren Wang, Ching-Chun Huang, Wei-Chen Chiu

Recently, diffusion-based blind super-resolution (SR) methods have shown great ability to generate high-resolution images with abundant high-frequency detail, but the detail is often achieved at the expense of fidelity. …

Blind Super-ResolutionSuper-Resolution

High-Resolution Image Synthesis with Latent Diffusion Models

2021-12-20 · CVPR 2022 1 · Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 외

By decomposing the image formation process into a sequential application of denoising autoencoders, diffusion models (DMs) achieve state-of-the-art synthesis results on image data and beyond. Additionally, their formulat…

DenoisingGPUImage GenerationImage Inpainting+6