Binary Document Image Super Resolution for Improved Readability and OCR Performance
There is a need for information retrieval from large collections of low-resolution (LR) binary document images, which can be found in digital libraries across the world, where the high-resolution (HR) counterpart is not available. This gives rise to the problem of binary document image super-resolution (BDISR). The objective of this paper is to address the interesting and challenging problem of super resolution of binary Tamil document images for improved readability and better optical character recognition (OCR). We propose multiple deep neural network architectures to address this problem and analyze their performance. The proposed models are all single image super-resolution techniques, which learn a generalized spatial correspondence between the LR and HR binary document images. We employ convolutional layers for feature extraction followed by transposed convolution and sub-pixel convolution layers for upscaling the features. Since the outputs of the neural networks are gray scale, we utilize the advantage of power law transformation as a post-processing technique to improve the character level pixel connectivity. The performance of our models is evaluated by comparing the OCR accuracies and the mean opinion scores given by human evaluators on LR images and the corresponding model-generated HR images.
Code (1)
Tasks
Image Super-ResolutionInformation RetrievalOptical Character RecognitionOptical Character Recognition (OCR)RetrievalSuper-ResolutionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Language Independent Single Document Image Super-Resolution using CNN for improved recognition
Recognition of document images have important applications in restoring old and classical texts. The problem involves quality improvement before passing it to a properly trained OCR to get accurate recognition of the tex…
Image EnhancementImage Super-ResolutionOptical Character Recognition (OCR)Super-ResolutionResults of improved fractional/integer order PDE-based binarization model
In this report, we present and compare the results of an improved fractional and integer order partial differential equation (PDE)-based binarization scheme. The improved model incorporates a diffusion term in addition t…
BinarizationEdge DetectionJoint Learning of Blind Super-Resolution and Crack Segmentation for Realistic Degraded Images
This paper proposes crack segmentation augmented by super resolution (SR) with deep neural networks. In the proposed method, a SR network is jointly trained with a binary segmentation network in an end-to-end manner. Thi…
Blind Super-ResolutionCrack SegmentationSegmentationSuper-ResolutionCascaded Detail-Preserving Networks for Super-Resolution of Document Images
The accuracy of OCR is usually affected by the quality of the input document image and different kinds of marred document images hamper the OCR results. Among these scenarios, the low-resolution image is a common and cha…
Image Super-ResolutionOptical Character Recognition (OCR)Super-ResolutionTask-driven single-image super-resolution reconstruction of document scans
Super-resolution reconstruction is aimed at generating images of high spatial resolution from low-resolution observations. State-of-the-art super-resolution techniques underpinned with deep learning allow for obtaining r…
Image Super-ResolutionOptical Character RecognitionSuper-ResolutionText Detection