paper-with-me

Papers

Printed Arabic Text Recognition using Linear and Nonlinear Regression

2017-02-05 · Ashraf A. Shahin

Arabic language is one of the most popular languages in the world. Hundreds of millions of people in many countries around the world speak Arabic as their native speaking. However, due to complexity of Arabic language, recognition of printed and handwritten Arabic text remained untouched for a very long time compared with English and Chinese. Although, in the last few years, significant number of researches has been done in recognizing printed and handwritten Arabic text, it stills an open research field due to cursive nature of Arabic script. This paper proposes automatic printed Arabic text recognition technique based on linear and ellipse regression techniques. After collecting all possible forms of each character, unique code is generated to represent each character form. Each code contains a sequence of lines and ellipses. To recognize fonts, a unique list of codes is identified to be used as a fingerprint of font. The proposed technique has been evaluated using over 14000 different Arabic words with different fonts and experimental results show that average recognition rate of the proposed technique is 86%.

📄 PDF Abstract BibTeX arXiv:1702.01444

Code (0)

등록된 구현이 없습니다.

Tasks

regression

Similar Papers 제목 키워드 기반

A Hybrid Deep Learning Model for Arabic Text Recognition

2020-09-04 · Mohammad Fasha, Bassam Hammo, Nadim Obeid, Jabir Widian

Arabic text recognition is a challenging task because of the cursive nature of Arabic writing system, its joint writing scheme, the large number of ligatures and many other challenges. Deep Learning DL models achieved si…

Deep LearningIrregular Text Recognition

Advancements and Challenges in Arabic Optical Character Recognition: A Comprehensive Survey

2023-12-19 · Mahmoud SalahEldin Kasem, Mohamed Mahmoud, Hyun-Soo Kang

Optical character recognition (OCR) is a vital process that involves the extraction of handwritten or printed text from scanned or printed images, converting it into a format that can be understood and processed by machi…

ArticlesOptical Character RecognitionOptical Character Recognition (OCR)

Arabic Character Segmentation Using Projection Based Approach with Profile's Amplitude Filter

2017-07-04 · Mahmoud A. A. Mousa, Mohammed S. Sayed, Mahmoud I. Abdalla

Arabic is one of the languages that present special challenges to Optical character recognition (OCR). The main challenge in Arabic is that it is mostly cursive. Therefore, a segmentation process must be carried out to d…

Optical Character RecognitionOptical Character Recognition (OCR)Segmentation

AraMS-28k: The Largest Publicly Released Line-Level Dataset of Historical Arabic Manuscripts with Margin and Insertion-Anchor Annotations

2026-08-27 · Mohamed Guechaoui, Mohamed Diaa Zellagui, Souleyman Chaib, Sahraoui Dhelim arxiv

We introduce AraMS-28k, the largest publicly released line-level dataset of genuine historical Arabic manuscripts, comprising 14 books, 3,043 pages, and 28,600 annotated text lines (27,971 main-text, 629 margin). Thirtee…

A Fuzzy Based Model to Identify Printed Sinhala Characters (ICIAfS14)

2014-12-24 · G. I. Gunarathna, M. A. P. Chamikara, R. G. Ragel

Character recognition techniques for printed documents are widely used for English language. However, the systems that are implemented to recognize Asian languages struggle to increase the accuracy of recognition. Among …