paper-with-me

홈 › Papers

Urdu text in natural scene images: a new dataset and preliminary text detection

2021-09-16 · Hazrat Ali, Khalid Iqbal, Ghulam Mujtaba, Ahmad Fayyaz, Mohammad Farhad Bulbul, Fazal Wahab Karam, Ali Zahir

Text detection in natural scene images for content analysis is an interesting task. The research community has seen some great developments for English/Mandarin text detection. However, Urdu text extraction in natural scene images is a task not well addressed. In this work, firstly, a new dataset is introduced for Urdu text in natural scene images. The dataset comprises of 500 standalone images acquired from real scenes. Secondly, the channel enhanced Maximally Stable Extremal Region (MSER) method is applied to extract Urdu text regions as candidates in an image. Two-stage filtering mechanism is applied to eliminate non-candidate regions. In the first stage, text and noise are classified based on their geometric properties. In the second stage, a support vector machine classifier is trained to discard non-text candidate regions. After this, text candidate regions are linked using centroid-based vertical and horizontal distances. Text lines are further analyzed by a different classifier based on HOG features to remove non-text regions. Extensive experimentation is performed on the locally developed dataset to evaluate the performance. The experimental results show good performance on test set images. The dataset will be made available for research use. To the best of our knowledge, the work is the first of its kind for the Urdu language and would provide a good dataset for free research use and serve as a baseline performance on the task of Urdu text extraction.

📄 PDF Abstract BibTeX arXiv:2109.08060

Code (0)

등록된 구현이 없습니다.

Tasks

Text Detection

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Dataset and Benchmark for Urdu Natural Scenes Text Detection, Recognition and Visual Question Answering

2024-05-21 · Hiba Maryam, Ling Fu, Jiajun Song, Tajrian ABM Shafayet 외

The development of Urdu scene text detection, recognition, and Visual Question Answering (VQA) technologies is crucial for advancing accessibility, information retrieval, and linguistic diversity in digital content, faci…

DiversityInformation RetrievalQuestion AnsweringRetrieval+4

Leveraging machine learning for less developed languages: Progress on Urdu text detection

2022-09-28 · Hazrat Ali

Text detection in natural scene images has applications for autonomous driving, navigation help for elderly and blind people. However, the research on Urdu text detection is usually hindered by lack of data resources. We…

Autonomous DrivingText Detection

A Benchmark Dataset and a Framework for Urdu Multimodal Named Entity Recognition

2025-05-08 · Hussain Ahmad, Qingyang Zeng, Jing Wan

The emergence of multimodal content, particularly text and images on social media, has positioned Multimodal Named Entity Recognition (MNER) as an increasingly important area of research within Natural Language Processin…

named-entity-recognitionNamed Entity Recognition

Transformer based Urdu Handwritten Text Optical Character Reader

2022-06-09 · Mohammad Daniyal Shaiq, Musa Dildar Ahmed Cheema, Ali Kamal

Extracting Handwritten text is one of the most important components of digitizing information and making it available for large scale setting. Handwriting Optical Character Reader (OCR) is a research problem in computer …

Natural Language UnderstandingOptical Character Recognition (OCR)Position

A Permuted Autoregressive Approach to Word-Level Recognition for Urdu Digital Text

2024-08-27 · Ahmed Mustafa, Muhammad Tahir Rafique, Muhammad Ijlal Baig, Hasan Sajid 외

This research paper introduces a novel word-level Optical Character Recognition (OCR) model specifically designed for digital Urdu text, leveraging transformer-based architectures and attention mechanisms to address the …

Data AugmentationOptical Character RecognitionOptical Character Recognition (OCR)