paper-with-me

홈 › Papers

Sequence to sequence learning for unconstrained scene text recognition

2016-07-20 · Ahmed Mamdouh A. Hassanien

In this work we present a state-of-the-art approach for unconstrained natural scene text recognition. We propose a cascade approach that incorporates a convolutional neural network (CNN) architecture followed by a long short term memory model (LSTM). The CNN learns visual features for the characters and uses them with a softmax layer to detect sequence of characters. While the CNN gives very good recognition results, it does not model relation between characters, hence gives rise to false positive and false negative cases (confusing characters due to visual similarities like "g" and "9", or confusing background patches with characters; either removing existing characters or adding non-existing ones) To alleviate these problems we leverage recent developments in LSTM architectures to encode contextual information. We show that the LSTM can dramatically reduce such errors and achieve state-of-the-art accuracy in the task of unconstrained natural scene text recognition. Moreover we manually remove all occurrences of the words that exist in the test set from our training set to test whether our approach will generalize to unseen data. We use the ICDAR 13 test set for evaluation and compare the results with the state of the art approaches [11, 18]. We finally present an application of the work in the domain of for traffic monitoring.

📄 PDF Abstract BibTeX arXiv:1607.06125

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Text Recognition

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Unconstrained Scene Text and Video Text Recognition for Arabic Script

2017-11-07 · Mohit Jain, Minesh Mathew, C. V. Jawahar

Building robust recognizers for Arabic has always been challenging. We demonstrate the effectiveness of an end-to-end trainable CNN-RNN hybrid architecture in recognizing Arabic text in videos and natural scenes. We outp…

Scene Text Recognition

Text Recognition in Real Scenarios with a Few Labeled Samples

2020-06-22 · Jinghuang Lin, Zhanzhan Cheng, Fan Bai, Yi Niu 외

Scene text recognition (STR) is still a hot research topic in computer vision field due to its various applications. Existing works mainly focus on learning a general model with a huge number of synthetic text images to …

Domain AdaptationScene Text Recognition

Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

2014-06-09 · Max Jaderberg, Karen Simonyan, Andrea Vedaldi, Andrew Zisserman

In this work we present a framework for the recognition of natural scene text. Our framework does not require any human-labelled data, and performs word recognition on the whole image holistically, departing from the cha…

Scene Text RecognitionText Generation

Unconstrained On-line Handwriting Recognition with Recurrent Neural Networks

2007-12-01 · NeurIPS 2007 12 · Alex Graves, Marcus Liwicki, Horst Bunke, Jürgen Schmidhuber 외

On-line handwriting recognition is unusual among sequence labelling tasks in that the underlying generator of the observed data, i.e. the movement of the pen, is recorded directly. However, the raw data can be difficult …

Handwriting RecognitionLanguage ModelingLanguage Modelling

An End-to-End Trainable Neural Network for Image-based Sequence Recognition and Its Application to Scene Text Recognition

2015-07-21 · Baoguang Shi, Xiang Bai, Cong Yao

Image-based sequence recognition has been a long-standing research topic in computer vision. In this paper, we investigate the problem of scene text recognition, which is among the most important and challenging tasks in…

Optical Character Recognition (OCR)Scene Text Recognition