paper-with-me

홈 › Papers

COCO-Text: Dataset and Benchmark for Text Detection and Recognition in Natural Images

2016-01-26 · Andreas Veit, Tomas Matera, Lukas Neumann, Jiri Matas, Serge Belongie

This paper describes the COCO-Text dataset. In recent years large-scale datasets like SUN and Imagenet drove the advancement of scene understanding and object recognition. The goal of COCO-Text is to advance state-of-the-art in text detection and recognition in natural images. The dataset is based on the MS COCO dataset, which contains images of complex everyday scenes. The images were not collected with text in mind and thus contain a broad variety of text instances. To reflect the diversity of text in natural scenes, we annotate text with (a) location in terms of a bounding box, (b) fine-grained classification into machine printed text and handwritten text, (c) classification into legible and illegible text, (d) script of the text and (e) transcriptions of legible text. The dataset contains over 173k text annotations in over 63k images. We provide a statistical analysis of the accuracy of our annotations. In addition, we present an analysis of three leading state-of-the-art photo Optical Character Recognition (OCR) approaches on our dataset. While scene text detection and recognition enjoys strong advances in recent years, we identify significant shortcomings motivating future work.

📄 PDF Abstract BibTeX arXiv:1601.07140

Code (4)

LinearPi/OCR_Chinese tf
OzHsu23/chineseocr tf
witcher425/CHINESEOCR tf
xiaofengShi/CHINESE-OCR tf

Tasks

DiversityGeneral ClassificationObject RecognitionOptical Character RecognitionOptical Character Recognition (OCR)Scene Text DetectionScene UnderstandingText Detection

Similar Papers 제목 키워드 기반

Human Keypoint Detection by Progressive Context Refinement

2019-10-27 · Jing Zhang, Zhe Chen, DaCheng Tao

Human keypoint detection from a single image is very challenging due to occlusion, blur, illumination and scale variance of person instances. In this paper, we find that context information plays an important role in add…

Human DetectionKeypoint DetectionMulti-Task Learning

Contextual Convolutional Neural Networks

2021-08-17 · Ionut Cosmin Duta, Mariana Iuliana Georgescu, Radu Tudor Ionescu

We propose contextual convolution (CoConv) for visual recognition. CoConv is a direct replacement of the standard convolution, which is the core component of convolutional neural networks. CoConv is implicitly equipped w…

Generative Adversarial Networkimage-classificationImage ClassificationImage Generation+2

Deep PCB To COCO Convertor

2022-05-01 · International Journal for Research in Engineering Application & Management (IJREAM) 2022 5 · Allena Venkata Sai Abhishek, Genti Sreeja, Dr K S Sudeep

Millions of datasets and many models use the input datasets in COCO format. In this paper, we are converting the Deep PCB dataset to COCO format. The Deep PCB is a manufacturing defect data set. It has 1500 image pairs. …

ClassificationData AugmentationData VisualizationDetecting Image Manipulation+15

BAN: Focusing on Boundary Context for Object Detection

2018-11-13 · Yonghyun Kim, Taewook Kim, Bong-Nam Kang, Jieun Kim 외

Visual context is one of the important clue for object detection and the context information for boundaries of an object is especially valuable. We propose a boundary aware network (BAN) designed to exploit the visual co…

Objectobject-detectionObject Detection

Benchmarking Object Detectors with COCO: A New Path Forward

2024-03-27 · Shweta Singh, Aayan Yadav, Jitesh Jain, Humphrey Shi 외

The Common Objects in Context (COCO) dataset has been instrumental in benchmarking object detectors over the past decade. Like every dataset, COCO contains subtle errors and imperfections stemming from its annotation pro…

BenchmarkingObjectobject-detectionObject Detection