paper-with-me

홈 › Papers

Mero Nagarikta: Advanced Nepali Citizenship Data Extractor with Deep Learning-Powered Text Detection and OCR

2024-10-08 · Sisir Dhakal, Sujan Sigdel, Sandesh Prasad Paudel, Sharad Kumar Ranabhat, Nabin Lamichhane

Transforming text-based identity documents, such as Nepali citizenship cards, into a structured digital format poses several challenges due to the distinct characteristics of the Nepali script and minor variations in print alignment and contrast across different cards. This work proposes a robust system using YOLOv8 for accurate text object detection and an OCR algorithm based on Optimized PyTesseract. The system, implemented within the context of a mobile application, allows for the automated extraction of important textual information from both the front and the back side of Nepali citizenship cards, including names, citizenship numbers, and dates of birth. The final YOLOv8 model was accurate, with a mean average precision of 99.1% for text detection on the front and 96.1% on the back. The tested PyTesseract optimized for Nepali characters outperformed the standard OCR regarding flexibility and accuracy, extracting text from images with clean and noisy backgrounds and various contrasts. Using preprocessing steps such as converting the images into grayscale, removing noise from the images, and detecting edges further improved the system's OCR accuracy, even for low-quality photos. This work expands the current body of research in multilingual OCR and document analysis, especially for low-resource languages such as Nepali. It emphasizes the effectiveness of combining the latest object detection framework with OCR models that have been fine-tuned for practical applications.

📄 PDF Abstract BibTeX arXiv:2410.05721

Code (0)

등록된 구현이 없습니다.

Tasks

object-detectionObject DetectionOptical Character Recognition (OCR)Text Detection

Methods 이 논문이 사용한 방법론

YOLOv8 설명 없음

Similar Papers 제목 키워드 기반

Benchmarking BERT-based Models for Sentence-level Topic Classification in Nepali Language

2026-02-27 · Nischal Karki, Bipesh Subedi, Prakash Poudyal, Rupak Raj Ghimire 외 arxiv

Transformer-based models such as BERT have significantly advanced Natural Language Processing (NLP) across many languages. However, Nepali, a low-resource language written in Devanagari script, remains relatively underex…

Generative AI for Named Entity Recognition in Low-Resource Language Nepali

2025-03-12 · Sameer Neupane, Jeevan Chapagain, Nobal B. Niraula, Diwa Koirala

Generative Artificial Intelligence (GenAI), particularly Large Language Models (LLMs), has significantly advanced Natural Language Processing (NLP) tasks, such as Named Entity Recognition (NER), which involves identifyin…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER

Optical Text Recognition in Nepali and Bengali: A Transformer-based Approach

2024-04-03 · S M Rakib Hasan, Aakar Dhakal, Md Humaion Kabir Mehedi, Annajiat Alim Rasel

Efforts on the research and development of OCR systems for Low-Resource Languages are relatively new. Low-resource languages have little training data available for training Machine Translation systems or other systems. …

DecoderMachine TranslationOptical Character Recognition (OCR)

Analyzing Race and Country of Citizenship Bias in Wikidata

2021-08-11 · Zaina Shaik, Filip Ilievski, Fred Morstatter

As an open and collaborative knowledge graph created by users and bots, it is possible that the knowledge in Wikidata is biased in regards to multiple factors such as gender, race, and country of citizenship. Previous wo…

Neural Machine Translation: Hindi-Nepali

2019-08-01 · WS 2019 8 · Sahinur Rahman Laskar, Partha Pakray, B, Sivaji yopadhyay

With the extensive use of Machine Translation (MT) technology, there is progressively interest in directly translating between pairs of similar languages. Because the main challenge is to overcome the limitation of avail…

Machine TranslationNMTTranslation