Character Proposal Network for Robust Text Extraction
Maximally stable extremal regions (MSER), which is a popular method to generate character proposals/candidates, has shown superior performance in scene text detection. However, the pixel-level operation limits its capability for handling some challenging cases (e.g., multiple connected characters, separated parts of one character and non-uniform illumination). To better tackle these cases, we design a character proposal network (CPN) by taking advantage of the high capacity and fast computing of fully convolutional network (FCN). Specifically, the network simultaneously predicts characterness scores and refines the corresponding locations. The characterness scores can be used for proposal ranking to reject non-character proposals and the refining process aims to obtain the more accurate locations. Furthermore, considering the situation that different characters have different aspect ratios, we propose a multi-template strategy, designing a refiner for each aspect ratio. The extensive experiments indicate our method achieves recall rates of 93.88%, 93.60% and 96.46% on ICDAR 2013, SVT and Chinese2k datasets respectively using less than 1000 proposals, demonstrating promising performance of our character proposal network.
Code (0)
등록된 구현이 없습니다.
Tasks
Scene Text DetectionText DetectionSimilar Papers 제목 키워드 기반
Object Proposals for Text Extraction in the Wild
Object Proposals is a recent computer vision technique receiving increasing interest from the research community. Its main objective is to generate a relatively small set of bounding box proposals that are most likely to…
ObjectEnd-to-End Temporal Relation Extraction in the Clinical Domain
Temporal relation extraction is an important task in the clinical domain, as it allows a better understanding of the temporal context of clinical events. In this paper, we present an end-to end temporal relation extracti…
Joint Entity and Relation ExtractionRelationRelation ExtractionTemporal Relation ExtractionHandwritten character recognition using some (anti)-diagonal structural features
In this paper, we present a methodology for off-line handwritten character recognition. The proposed methodology relies on a new feature extraction technique based on structural characteristics, histograms and profiles. …
Video Object Segmentation through Spatially Accurate and Temporally Dense Extraction of Primary Object Regions
In this paper, we propose a novel approach to extract primary object segments in videos in the 'object proposal' domain. The extracted primary object regions are then used to build object models for optimized video segme…
ObjectOptical Flow EstimationSemantic SegmentationVideo Object Segmentation+2Context-aware Proposal Network for Temporal Action Detection
This technical report presents our first place winning solution for temporal action detection task in CVPR-2022 AcitivityNet Challenge. The task aims to localize temporal boundaries of action instances with specific clas…
Action ClassificationAction Detection