paper-with-me

홈 › Papers

FOSNet: An End-to-End Trainable Deep Neural Network for Scene Recognition

2019-07-17 · Hongje Seong, Junhyuk Hyun, Euntai Kim

Scene recognition is an image recognition problem aimed at predicting the category of the place at which the image is taken. In this paper, a new scene recognition method using the convolutional neural network (CNN) is proposed. The proposed method is based on the fusion of the object and the scene information in the given image and the CNN framework is named as FOS (fusion of object and scene) Net. In addition, a new loss named scene coherence loss (SCL) is developed to train the FOSNet and to improve the scene recognition performance. The proposed SCL is based on the unique traits of the scene that the 'sceneness' spreads and the scene class does not change all over the image. The proposed FOSNet was experimented with three most popular scene recognition datasets, and their state-of-the-art performance is obtained in two sets: 60.14% on Places 2 and 90.37% on MIT indoor 67. The second highest performance of 77.28% is obtained on SUN 397.

📄 PDF Abstract BibTeX arXiv:1907.07570

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Recognition

Similar Papers 제목 키워드 기반

Mask TextSpotter: An End-to-End Trainable Neural Network for Spotting Text with Arbitrary Shapes

2018-07-06 · ECCV 2018 9 · Pengyuan Lyu, Minghui Liao, Cong Yao, Wenhao Wu 외

Recently, models based on deep neural networks have dominated the fields of scene text detection and recognition. In this paper, we investigate the problem of scene text spotting, which aims at simultaneous text detectio…

Scene Text DetectionSemantic SegmentationText DetectionText Spotting

An End-to-End Trainable Neural Network for Image-based Sequence Recognition and Its Application to Scene Text Recognition

2015-07-21 · Baoguang Shi, Xiang Bai, Cong Yao

Image-based sequence recognition has been a long-standing research topic in computer vision. In this paper, we investigate the problem of scene text recognition, which is among the most important and challenging tasks in…

Optical Character Recognition (OCR)Scene Text Recognition

Deep TextSpotter: An End-To-End Trainable Scene Text Localization and Recognition Framework

2017-10-01 · ICCV 2017 10 · Michal Busta, Lukas Neumann, Jiri Matas

A method for scene text localization and recognition is proposed. The novelties include: training of both text detection and recognition in a single end-to-end pass, the structure of the recognition CNN and the geometry …

GPUText Detection

Scene Text Recognition With Finer Grid Rectification

2020-01-26 · Gang Wang

Scene Text Recognition is a challenging problem because of irregular styles and various distortions. This paper proposed an end-to-end trainable model consists of a finer rectification module and a bidirectional attentio…

DecoderScene Text Recognition

Mask TextSpotter: An End-to-End Trainable Neural Network for Spotting Text with Arbitrary Shapes

2019-08-22 · ECCV 2018 9 · Minghui Liao, Pengyuan Lyu, Minghang He, Cong Yao 외

Unifying text detection and text recognition in an end-to-end training fashion has become a new trend for reading text in the wild, as these two tasks are highly relevant and complementary. In this paper, we investigate …

Scene Text RecognitionSemantic SegmentationText DetectionText Spotting