PlugNet: Degradation Aware Scene Text Recognition Supervised by a Pluggable Super-Resolution Unit
In this paper, we address the problem of recognizing degradation images that are suffering from high blur or low-resolution. We propose a novel degradation aware scene text recognizer with a pluggable super-resolution unit (PlugNet) to recognize low-quality scene text to solve this task from the feature-level. The whole networks can be trained end-to-end with a pluggable super-resolution unit (PSU) and the PSU will be removed after training so that it brings no extra computation. The PSU aims to obtain a more robust feature representation for recognizing low-quality text images. Moreover, to further improve the feature quality, we introduce two types of feature enhancement strategies: Feature Squeeze Module (FSM) which aims to reduce the loss of spatial acuity and Feature Enhance Module (FEM) which combines the feature maps from low to high to provide diversity semantics. As a consequence, the PlugNet achieves state-of-the-art performance on various widely used text recognition benchmarks like IIIT5K, SVT, SVTP, ICDAR15 and etc.
Code (1)
Tasks
DiversityMulti-Task LearningScene Text RecognitionSuper-ResolutionSimilar Papers 제목 키워드 기반
Scene Text Image Super-Resolution via Content Perceptual Loss and Criss-Cross Transformer Blocks
Text image super-resolution is a unique and important task to enhance readability of text images to humans. It is widely used as pre-processing in scene text recognition. However, due to the complex degradation in natura…
Image ReconstructionImage Super-ResolutionScene Text RecognitionSuper-ResolutionLearning a Unified Degradation-aware Representation Model for Multi-modal Image Fusion
All-in-One Degradation-Aware Fusion Models (ADFMs), a class of multi-modal image fusion models, address complex scenes by mitigating degradations from source images and generating high-quality fused images. Mainstream AD…
Infrared And Visible Image FusionScene Text Image Super-Resolution via Parallelly Contextual Attention Network
Optical degradation makes text shapes and edges blurred, so the existing scene text recognition methods are difficult to achieve desirable results on low-resolution (LR) scene text images acquired in natural scenes. Ther…
Image ReconstructionImage Super-ResolutionScene Text RecognitionSuper-ResolutionText-DIAE: A Self-Supervised Degradation Invariant Autoencoders for Text Recognition and Document Enhancement
In this paper, we propose a Text-Degradation Invariant Auto Encoder (Text-DIAE), a self-supervised model designed to tackle two tasks, text recognition (handwritten or scene-text) and document image enhancement. We start…
Document EnhancementImage EnhancementScene Text RecognitionNight-to-Day Translation via Illumination Degradation Disentanglement
Night-to-Day translation (Night2Day) aims to achieve day-like vision for nighttime scenes. However, processing night images with complex degradations remains a significant challenge under unpaired conditions. Previous me…
Contrastive LearningDisentanglementTranslation