paper-with-me

홈 › Papers

IndicSTR12: A Dataset for Indic Scene Text Recognition

2024-03-12 · Harsh Lunia, Ajoy Mondal, C V Jawahar

The importance of Scene Text Recognition (STR) in today's increasingly digital world cannot be overstated. Given the significance of STR, data intensive deep learning approaches that auto-learn feature mappings have primarily driven the development of STR solutions. Several benchmark datasets and substantial work on deep learning models are available for Latin languages to meet this need. On more complex, syntactically and semantically, Indian languages spoken and read by 1.3 billion people, there is less work and datasets available. This paper aims to address the Indian space's lack of a comprehensive dataset by proposing the largest and most comprehensive real dataset - IndicSTR12 - and benchmarking STR performance on 12 major Indian languages. A few works have addressed the same issue, but to the best of our knowledge, they focused on a small number of Indian languages. The size and complexity of the proposed dataset are comparable to those of existing Latin contemporaries, while its multilingualism will catalyse the development of robust text detection and recognition models. It was created specifically for a group of related languages with different scripts. The dataset contains over 27000 word-images gathered from various natural scenes, with over 1000 word-images for each language. Unlike previous datasets, the images cover a broader range of realistic conditions, including blur, illumination changes, occlusion, non-iconic texts, low resolution, perspective text etc. Along with the new dataset, we provide a high-performing baseline on three models - PARSeq, CRNN, and STARNet.

📄 PDF Abstract BibTeX arXiv:2403.08007

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingScene Text RecognitionText Detection

Similar Papers 제목 키워드 기반

Reading in the Dark: Low-light Scene Text Recognition

2026-04-26 · Xuanshuo Fu, Lei Kang, Ernest Valveny, Dimosthenis Karatzas 외 arxiv

Accurate text recognition in low-light environments is essential for intelligent systems in applications ranging from autonomous vehicles to smart surveillance. However, challenges such as poor illumination and noise int…

Low-Light Image EnhancementScene Text RecognitionAutonomous Vehicles

Semantic-Aware Scene Recognition

2019-09-05 · Alejandro López-Cifuentes, Marcos Escudero-Viñolo, Jesús Bescós, Álvaro García-Martín

Scene recognition is currently one of the top-challenging research fields in computer vision. This may be due to the ambiguity between classes: images of several scene classes may share similar objects, which causes conf…

Scene ClassificationScene RecognitionSemantic Segmentation

CMFN: Cross-Modal Fusion Network for Irregular Scene Text Recognition

2024-01-18 · Jinzhi Zheng, Ruyi Ji, Libo Zhang, Yanjun Wu 외

Scene text recognition, as a cross-modal task involving vision and text, is an important research topic in computer vision. Most existing methods use language models to extract semantic information for optimizing visual …

PositionScene Text Recognition

Benchmarking Scene Text Recognition in Devanagari, Telugu and Malayalam

2021-04-09 · Minesh Mathew, Mohit Jain, CV Jawahar

Inspired by the success of Deep Learning based approaches to English scene text recognition, we pose and benchmark scene text recognition for three Indic scripts - Devanagari, Telugu and Malayalam. Synthetic word images …

BenchmarkingScene Text Recognition

On Manipulating Scene Text in the Wild with Diffusion Models

2023-11-01 · Joshua Santoso, Christian Simon, Williem Pao

Diffusion models have gained attention for image editing yielding impressive results in text-to-image tasks. On the downside, one might notice that generated images of stable diffusion models suffer from deteriorated det…

Optical Character RecognitionOptical Character Recognition (OCR)Scene Text Editing