Incidental Scene Text Understanding: Recent Progresses on ICDAR 2015 Robust Reading Competition Challenge 4
Different from focused texts present in natural images, which are captured with user's intention and intervention, incidental texts usually exhibit much more diversity, variability and complexity, thus posing significant difficulties and challenges for scene text detection and recognition algorithms. The ICDAR 2015 Robust Reading Competition Challenge 4 was launched to assess the performance of existing scene text detection and recognition methods on incidental texts as well as to stimulate novel ideas and solutions. This report is dedicated to briefly introduce our strategies for this challenging problem and compare them with prior arts in this field.
Code (0)
등록된 구현이 없습니다.
Tasks
DiversityScene Text DetectionText DetectionSimilar Papers 제목 키워드 기반
Deep Direct Regression for Multi-Oriented Scene Text Detection
In this paper, we first provide a new perspective to divide existing high performance object detection methods into direct and indirect regressions. Direct regression performs boundary regression by predicting the offset…
Multi-Oriented Scene Text Detectionobject-detectionObject Detectionregression+2Semantic Stereo for Incidental Satellite Images
The increasingly common use of incidental satellite images for stereo reconstruction versus rigidly tasked binocular or trinocular coincident collection is helping to enable timely global-scale 3D mapping; however, relia…
3D ReconstructionScene SegmentationSegmentationSoft-PHOC Descriptor for End-to-End Word Spotting in Egocentric Scene Images
Word spotting in natural scene images has many applications in scene understanding and visual assistance. In this paper we propose a technique to create and exploit an intermediate representation of images based on text …
AttributeDynamic Time WarpingScene UnderstandingOn the General Value of Evidence, and Bilingual Scene-Text Visual Question Answering
Visual Question Answering (VQA) methods have made incredible progress, but suffer from a failure to generalize. This is visible in the fact that they are vulnerable to learning coincidental correlations in the data rathe…
Question AnsweringReferring ExpressionVisual Question AnsweringVisual Question Answering (VQA)Deep Matching Prior Network: Toward Tighter Multi-oriented Text Detection
Detecting incidental scene text is a challenging task because of multi-orientation, perspective distortion, and variation of text size, color and scale. Retrospective research has only focused on using rectangular boundi…
Text Detection