paper-with-me

홈 › Papers

Font Acknowledgment and Character Extraction of Digital and Scanned Images

2013-05-17 · Syed Muhammad Arsalan Bashir

The font recognition and character extraction is of immense importance as these are many scenarios where data are in such a form, which cannot be processed like in image form or as a hard copy. So the procedure developed in this paper is basically related to identifying the font (Times New Roman, Arial and Comic Sans MS) and afterwards recovering the text using simple correlation based method where the binary templates are correlated to the input image text characters. All of this extraction is done in the presence of a little noise as images may have noisy patterns due to photocopying. The significance of this method exists in extraction of data from various monitoring (Surveillance) camera footages or even more. The method is developed on Matlab\c{opyright} which takes input image and recovers text and font information from it in a text file.

📄 PDF Abstract BibTeX arXiv:1305.4064

Code (0)

등록된 구현이 없습니다.

Tasks

Font Recognition

Similar Papers 제목 키워드 기반

An Efficient Language-Independent Multi-Font OCR for Arabic Script

2020-09-18 · Hussein Osman, Karim Zaghw, Mostafa Hazem, Seifeldin Elsehely

Optical Character Recognition (OCR) is the process of extracting digitized text from images of scanned documents. While OCR systems have already matured in many languages, they still have shortcomings in cursive language…

Optical Character RecognitionOptical Character Recognition (OCR)Segmentation

Optimizing Nepali PDF Extraction: A Comparative Study of Parser and OCR Technologies

2024-07-05 · Prabin Paudel, Supriya Khadka, Ranju G. C., Rahul Shah

This research compares PDF parsing and Optical Character Recognition (OCR) methods for extracting Nepali content from PDFs. PDF parsing offers fast and accurate extraction but faces challenges with non-Unicode Nepali fon…

Optical Character RecognitionOptical Character Recognition (OCR)

ScanBank: A Benchmark Dataset for Figure Extraction from Scanned Electronic Theses and Dissertations

2021-06-23 · Sampanna Yashwant Kahu, William A. Ingram, Edward A. Fox, Jian Wu

We focus on electronic theses and dissertations (ETDs), aiming to improve access and expand their utility, since more than 6 million are publicly available, and they constitute an important corpus to aid research and edu…

Data AugmentationTable Extraction

MX-Font++: Mixture of Heterogeneous Aggregation Experts for Few-shot Font Generation

2025-03-04 · Weihang Wang, Duolin Sun, Jielei Zhang, Longwen Gao

Few-shot Font Generation (FFG) aims to create new font libraries using limited reference glyphs, with crucial applications in digital accessibility and equity for low-resource languages, especially in multilingual artifi…

Font GenerationMixture-of-Experts

Omnifont Persian OCR System Using Primitives

2022-02-13 · Azarakhsh Keipour, Mohammad Eshghi, Sina Mohammadzadeh Ghadikolaei, Negin Mohammadi 외

In this paper, we introduce a model-based omnifont Persian OCR system. The system uses a set of 8 primitive elements as structural features for recognition. First, the scanned document is preprocessed. After normalizing …

Optical Character Recognition (OCR)