SuperOCR: A Conversion from Optical Character Recognition to Image Captioning
Optical Character Recognition (OCR) has many real world applications. The existing methods normally detect where the characters are, and then recognize the character for each detected location. Thus the accuracy of characters recognition is impacted by the performance of characters detection. In this paper, we propose a method for recognizing characters without detecting the location of each character. This is done by converting the OCR task into an image captioning task. One advantage of the proposed method is that the labeled bounding boxes for the characters are not needed during training. The experimental results show the proposed method outperforms the existing methods on both the license plate recognition and the watermeter character recognition tasks. The proposed method is also deployed into a low-power (300mW) CNN accelerator chip connected to a Raspberry Pi 3 for on-device applications.
Code (0)
등록된 구현이 없습니다.
Tasks
Image CaptioningLicense Plate RecognitionOptical Character RecognitionOptical Character Recognition (OCR)Raspberry Pi 3Similar Papers 제목 키워드 기반
Confronting the Constraints for Optical Character Segmentation from Printed Bangla Text Image
In a world of digitization, optical character recognition holds the automation to written history. Optical character recognition system basically converts printed images into editable texts for better storage and usabili…
Optical Character RecognitionOptical Character Recognition (OCR)SegmentationMultiple-image encryption and hiding with an optical diffractive neural network
A cascaded phase-only mask architecture (or an optical diffractive neural network) can be employed for different optical information processing tasks such as pattern recognition, orbital angular momentum (OAM) mode conve…
10-shot image generationOCR accuracy improvement on document images through a novel pre-processing approach
Digital camera and mobile document image acquisition are new trends arising in the world of Optical Character Recognition and text detection. In some cases, such process integrates many distortions and produces poorly sc…
BinarizationOptical Character RecognitionOptical Character Recognition (OCR)Text DetectionSuperOCR for ALTA 2017 Shared Task
Artificial Eye for the Blind
The main backbone of our Artificial Eye model is the Raspberry pi3 which is connected to the webcam ,ultrasonic proximity sensor, speaker and we also run all our software models i.e object detection, Optical Character re…
Objectobject-detectionObject DetectionOptical Character Recognition+3