paper-with-me

홈 › Papers

Efficient scene text image super-resolution with semantic guidance

2024-03-20 · LeoWu TomyEnrique, Xiangcheng Du, Kangliang Liu, Han Yuan, Zhao Zhou, Cheng Jin

Scene text image super-resolution has significantly improved the accuracy of scene text recognition. However, many existing methods emphasize performance over efficiency and ignore the practical need for lightweight solutions in deployment scenarios. Faced with the issues, our work proposes an efficient framework called SGENet to facilitate deployment on resource-limited platforms. SGENet contains two branches: super-resolution branch and semantic guidance branch. We apply a lightweight pre-trained recognizer as a semantic extractor to enhance the understanding of text information. Meanwhile, we design the visual-semantic alignment module to achieve bidirectional alignment between image features and semantics, resulting in the generation of highquality prior guidance. We conduct extensive experiments on benchmark dataset, and the proposed SGENet achieves excellent performance with fewer computational costs. Code is available at https://github.com/SijieLiu518/SGENet

📄 PDF Abstract BibTeX arXiv:2403.13330

Code (1)

sijieliu518/sgenet 공식 구현 pytorch

Tasks

Image Super-ResolutionScene Text RecognitionSuper-Resolution

Similar Papers 제목 키워드 기반

Recognition-Guided Diffusion Model for Scene Text Image Super-Resolution

2023-11-22 · Yuxuan Zhou, Liangcai Gao, Zhi Tang, Baole Wei

Scene Text Image Super-Resolution (STISR) aims to enhance the resolution and legibility of text within low-resolution (LR) images, consequently elevating recognition accuracy in Scene Text Recognition (STR). Previous met…

DenoisingDiversityImage Super-ResolutionScene Text Recognition+1

PEAN: A Diffusion-Based Prior-Enhanced Attention Network for Scene Text Image Super-Resolution

2023-11-29 · Zuoyan Zhao, Hui Xue, Pengfei Fang, Shipeng Zhu

Scene text image super-resolution (STISR) aims at simultaneously increasing the resolution and readability of low-resolution scene text images, thus boosting the performance of the downstream recognition task. Two factor…

Image Super-ResolutionMulti-Task LearningSuper-Resolution

Multi-Resolution Alignment for Voxel Sparsity in Camera-Based 3D Semantic Scene Completion

2026-02-03 · Zhiwen Yang, Yuxin Peng arxiv

Camera-based 3D semantic scene completion (SSC) offers a cost-effective solution for assessing the geometric occupancy and semantic labels of each voxel in the surrounding 3D scene with image inputs, providing a voxel-le…

3D Semantic Scene CompletionAutonomous Driving

Towards Robust Scene Text Image Super-resolution via Explicit Location Enhancement

2023-07-19 · Hang Guo, Tao Dai, Guanghao Meng, Shu-Tao Xia

Scene text image super-resolution (STISR), aiming to improve image quality while boosting downstream scene text recognition accuracy, has recently achieved great success. However, most existing methods treat the foregrou…

Image Super-ResolutionLEMMAScene Text RecognitionSuper-Resolution

Improving Scene Text Image Super-resolution via Dual Prior Modulation Network

2023-02-21 · Shipeng Zhu, Zuoyan Zhao, Pengfei Fang, Hui Xue

Scene text image super-resolution (STISR) aims to simultaneously increase the resolution and legibility of the text images, and the resulting images will significantly affect the performance of downstream tasks. Although…

Image Super-ResolutionSuper-Resolution